影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
90.91/100
前 0.5%
全站排名 #333
发表论文29 篇
平均评分
年均产出9.7 篇/年
Lijuan Wang
研究方向
computer vision · vision and language · multi-modal
25
EdiVal-Agent: An Object-Centric Framework for Automated, Fine-Grained Evaluation of Multi-Turn Editing
ICLR 2026Poster
26
STITCH: Simultaneous Thinking and Talking with Chunked Reasoning for Spoken Language Models
ICLR 2026Poster
通讯10
TextAtlas5M: A Large-Scale Dataset for Long and Structured Text Image Generation
ICLR 2026Rejected
22
InfoAgent: Advancing Autonomous Information‑Seeking Agents
ICLR 2026Rejected
22
V-MAGE: A Game Evaluation Framework for Assessing Vision-Centric Capabilities in Multimodal Large Language Models
ICLR 2026Withdrawn
通讯5
Where do Reasoning Models Make a Difference? Follow the Reasoning Leader for Efficient Decoding
ICLR 2026Withdrawn
通讯6
MV-Diffus3R: Refining Multi-View Diffusions for Geometric Coherence 3D Reconstruction
ICLR 2026Rejected
通讯24
Shanks: Simultaneous Hearing and Thinking for Spoken Language Models
ICLR 2026Withdrawn
通讯4
Computer-Use Agents as Judges for Automatic GUI Design
ICLR 2026Withdrawn
6
The Agent's Marathon: Probing the Limits of Endurance in Long-Horizon Tasks
ICLR 2026Rejected
11
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
ICML 2025Oral
15
MMIE: Massive Multimodal Interleaved Comprehension Benchmark for Large Vision-Language Models
ICLR 2025Oral
24
SlowFast-VGen: Slow-Fast Learning for Action-Driven Long Video Generation
ICLR 2025Spotlight
通讯21
VAGEN: Reinforcing World Model Reasoning for Multi-Turn VLM Agents
NeurIPS 2025Poster
22
SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
NeurIPS 2025Spotlight
通讯24
ViCrit: A Verifiable Reinforcement Learning Proxy Task for Visual Perception in VLMs
NeurIPS 2025Poster
通讯20
EditRoom: LLM-parameterized Graph Diffusion for Composable 3D Room Layout Editing
ICLR 2025Poster
19
Tuning Timestep-Distilled Diffusion Model Using Pairwise Sample Optimization
ICLR 2025Poster
26
Point-RFT: Improving Multimodal Reasoning with Visually Grounded Reinforcement Finetuning
NeurIPS 2025Poster
通讯6
GenXD: Generating Any 3D and 4D Scenes
ICLR 2025Poster
通讯18
CertainlyUncertain: A Benchmark and Metric for Multimodal Epistemic and Aleatoric Awareness
ICLR 2025Poster
39
MMWorld: Towards Multi-discipline Multi-faceted World Model Evaluation in Videos
ICLR 2025Poster