影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
68.24/100
前 2.9%
全站排名 #1,887
发表论文16 篇
平均评分
年均产出8.0 篇/年
Jiahao Xu
研究方向
Sentence Embeddings · Machine Translation · Reinforcement Learning from Human Feedback · Large Language Models · LLM Reasoning
13
WebDevJudge: Evaluating (M)LLMs as Critiques for Web Development Quality
ICLR 2026Oral
20
DeepMath-103K: A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning
ICLR 2026Poster
三作15
THE END OF MANUAL DECODING: TOWARDS TRULY END-TO-END LANGUAGE MODELS
ICLR 2026Poster
31
Learning to Reward: A Contextual Bandit Framework for Distributional Reward Policy Optimization
ICLR 2026Rejected
24
DeepCompress: A Dual Reward Strategy for Dynamically Exploring and Compressing Reasoning Chains
ICLR 2026Poster
16
DeepTheorem: Advancing LLM Reasoning for Theorem Proving Through Natural Language and Reinforcement Learning
ICLR 2026Rejected
二作6
Less is More: Denoising Knowledge Graphs For Retrieval Augmented Generation
ICLR 2026Rejected
18
Thoughts Are All Over the Place: On the Underthinking of Long Reasoning Models
NeurIPS 2025Spotlight
三作19
RaSA: Rank-Sharing Low-Rank Adaptation
ICLR 2025Poster
28
Trust, But Verify: A Self-Verification Approach to Reinforcement Learning with Verifiable Rewards
NeurIPS 2025Poster
23
Two Experts Are All You Need for Steering Thinking: Reinforcing Cognitive Effort in MoE Reasoning Models Without Additional Training
NeurIPS 2025Poster
9
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM’s Reasoning Capability
ICML 2025Poster
三作26
The First Few Tokens Are All You Need: An Efficient and Effective Unsupervised Prefix Fine-Tuning Method for Reasoning Models
NeurIPS 2025Poster
二作11
Do NOT Think That Much for 2+3=? On the Overthinking of Long Reasoning Models
ICML 2025Poster
二作11
Thoughts Are All Over the Place: On the Underthinking of o1-Like LLMs
ICML 2025Rejected
三作6
Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training
ICLR 2025Withdrawn