影响力指数
90.01/100
前 0.6%
全站排名 #373
发表论文41
平均评分5.2
年均产出13.7 篇/年

Xiaodan Liang

Full Professor@SUN YAT-SEN UNIVERSITY·中国·OpenReview
研究方向

Embodied Vision · Cross-modal Understanding and generation · Image/Video Generation and Editing

7.1
24

SPC: Evolving Self-Play Critic via Adversarial Games for LLM Reasoning

NeurIPS 2025Poster
7.0
17

MMTryon: Multi-Modal Multi-Reference Control for High-Quality Fashion Generation

ICLR 2025Rejected
通讯
6.8
31

WISA: World simulator assistant for physics-aware text-to-video generation

NeurIPS 2025Spotlight
通讯
6.7
25

OptiBench Meets ReSocratic: Measure and Improve LLMs for Optimization Modeling

ICLR 2025Poster
6.4
10

GDrag:Towards General-Purpose Interactive Editing with Anti-ambiguity Point Diffusion

ICLR 2025Poster
通讯
6.4
26

PT-T2I/V: An Efficient Proxy-Tokenized Diffusion Transformer for Text-to-Image/Video-Task

ICLR 2025Poster
通讯
6.3
12

CatVTON: Concatenation Is All You Need for Virtual Try-On with Diffusion Models

ICLR 2025Poster
通讯
6.0
29

UniGS: Unified Language-Image-3D Pretraining with Gaussian Splatting

ICLR 2025Poster
通讯
5.7
23

Sitcom-Crafter: A Plot-Driven Human Motion Generation System in 3D Scenes

ICLR 2025Poster
通讯
5.5
13

S2-Track: A Simple yet Strong Approach for End-to-End 3D Multi-Object Tracking

ICML 2025Poster
通讯
5.3
5

EMOVA: Empowering Language Models to See, Hear and Speak with Vivid Emotions

ICLR 2025Withdrawn
4.8
30

UncertaintyRAG: Span Uncertainty Enhanced Long-Context Modeling for Retrieval-Augmented Generation

ICLR 2025Rejected
4.8
5

Continual LLaVA: Continual Instruction Tuning in Large Vision-Language Models

ICLR 2025Withdrawn
通讯
4.8
5

HiRes-LLaVA: Restoring Fragmentation Input in High-Resolution Large Vision-Language Models

ICLR 2025Withdrawn
通讯
4.7
5

StoryAgent: Customized Storytelling Video Generation via Multi-Agent Collaboration

ICLR 2025Rejected
通讯
4.0
5

Memory-Driven Multimodal Chain of Thought for Embodied Long-Horizon Task Planning

ICLR 2025Withdrawn
通讯
3.0
6

ActionFiller: Fill-In-The-Blank Prompting for OS Agent

ICLR 2025Rejected
通讯