影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
64.45/100
前 3.7%
全站排名 #2,371
发表论文16 篇
平均评分
年均产出8.0 篇/年
Xiaoyu Tan
研究方向
Large Language Models · Reinforcement Learning
17
PRISM: Festina Lente Proactivity—Risk-Sensitive, Uncertainty-Aware Deliberation for Proactive Agents
ICLR 2026Poster
二作30
The Choice of Divergence: A Neglected Key to Mitigating Diversity Collapse in Reinforcement Learning with Verifiable Reward
ICLR 2026Poster
30
Count Counts: Motivating Exploration in LLM Reasoning with Count-based Intrinsic Rewards
ICLR 2026Poster
20
CUARewardBench: Benchmark for Evaluating Reward Models on Computer-using Agent Trajectories
ICLR 2026Rejected
二作21
Learn the Ropes, Then Trust the Wins: Self-imitation with Progressive Exploration for Agentic Reinforcement Learning
ICLR 2026Poster
二作6
Asleep at the Wheel: Benchmarking the Inattention of Vision Language Models to Clinical Sleep Signals
ICLR 2026Rejected
通讯27
Distributional Reinforcement Learning for Large Language Models
ICLR 2026Rejected
15
Training-Free Group Relative Policy Optimization
ICLR 2026Rejected
4
Transitioning from Full-Context to Active Evidence-Seeking Evaluation: A Novel Benchmark for Real-World Artificial Intelligence Assisted Medical Diagnosis
ICLR 2026Withdrawn
二作4
Dual-Route Mental Imagery for Robust VLM-based Medical Image Diagnosis
ICLR 2026Withdrawn
二作5
RoRecomp: Enhancing Reasoning Efficiency via Rollout Response Recomposition in Reinforcement Learning
ICLR 2026Rejected
三作6
MMUDA: Towards Robust Sleep Staging with Multi-Source Multi-Channel Unsupervised Domain Adaptation
ICLR 2026Rejected
三作21
ORIGAMISPACE: Benchmarking Multimodal LLMs in Multi-Step Spatial Reasoning with Mathematical Constraints
NeurIPS 2025Spotlight
22
Atomic Thinking of LLMs: Decoupling and Exploring Mathematical Reasoning Abilities
NeurIPS 2025Poster
19
Refine Knowledge of Large Language Models via Adaptive Contrastive Learning
ICLR 2025Poster
15
One Example Shown, Many Concepts Known! Counterexample-Driven Conceptual Reasoning in Mathematical LLMs
ICML 2025Poster