影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
50.87/100
前 8.2%
全站排名 #5,255
发表论文11 篇
平均评分
年均产出3.7 篇/年
21
Pretrain Value, Not Reward: Decoupled Value Policy Optimization
ICLR 2026Poster
19
RePrompt: Reasoning-Augmented Reprompting for Text-to-Image Generation via Reinforcement Learning
ICLR 2026Poster
三作16
WarriorMath: Empowering Mathematical Reasoning for Large Language Models via Expert Battles
ICLR 2026Rejected
14
Learning to Refine: Self-Refinement of Parallel Reasoning in LLMs
ICLR 2026Withdrawn
二作5
Duet: Joint Exploration of User–Item Profiles
ICLR 2026Withdrawn
30
WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct
ICLR 2025Oral
37
SELF-EVOLVED REWARD LEARNING FOR LLMS
ICLR 2025Poster
29
Thread: A Logic-Based Data Organization Paradigm for How-To Question Answering with Retrieval Augmented Generation
ICLR 2025Rejected