影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
82.02/100
前 1.2%
全站排名 #772
发表论文25 篇
平均评分
年均产出8.3 篇/年
Xiaoye Qu
研究方向
Multimodality and Language Grounding · Information extraction
22
Spotlight on Token Perception for Multimodal Reinforcement Learning
ICLR 2026Poster
二作21
ExGRPO: Learning to Reason from Experience
ICLR 2026Poster
26
FrameThinker: Learning to Think with Long Videos via Multi-Turn Frame Spotlighting
ICLR 2026Poster
二作24
Diversity-Incentivized Exploration for Versatile Reasoning
ICLR 2026Poster
27
Towards an AI Musician: Synthesizing Sheet Music Problems for Musical Reasoning
ICLR 2026Rejected
5
Flash-DMD: Unifying Distillation and Refinement for High-Fidelity Few-Step Image Generation
ICLR 2026Withdrawn
19
Revisual-R1: Advancing Multimodal Reasoning From Optimized Cold Start to Staged Reinforcement Learning
ICLR 2026Poster
5
Scaling Reasoning, Losing Control: Evaluating Instruction Following in Large Reasoning Models
ICLR 2026Withdrawn
14
Learning to Reason under Off-Policy Guidance
NeurIPS 2025Poster
23
Towards Building Model/Prompt-Transferable Attackers against Large Vision-Language Models
NeurIPS 2025Spotlight
三作23
Fit the Distribution: Cross-Image/Prompt Adversarial Attacks on Multimodal Large Language Models
NeurIPS 2025Poster
17
Divide and Conquer: Grounding LLMs as Efficient Decision-Making Agents via Offline Hierarchical Reinforcement Learning
ICML 2025Poster
三作9
Make LoRA Great Again: Boosting LoRA with Adaptive Singular Values and Mixture-of-Experts Optimization Alignment
ICML 2025Poster
6
LASP-2: Rethinking Sequence Parallelism for Linear Attention and its Hybrid
ICLR 2025Rejected
三作20
CLIP-MoE: Towards Building Mixture of Experts for CLIP with Diversified Multiplet Upcycling
ICLR 2025Rejected
三作13
Does Your Video-language Model Actually Understand the Language Input?
ICLR 2025Withdrawn
通讯11
Test-Time Preference Optimization: On-the-Fly Alignment via Iterative Textual Feedback
ICML 2025Poster
三作6
Demystifying the Underappreciated Long-Tail Problems in Large Vision Language Models
ICLR 2025Withdrawn
二作6
Strategy-centric Synthesis: Connecting Billions of Image-Text Pairs to High-Quality Visual Instruction Data
ICLR 2025Withdrawn
5
Are Large Vision-Language Models Robust to Adversarial Visual Transformations?
ICLR 2025Withdrawn