影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
50.99/100
前 8.1%
全站排名 #5,214
发表论文12 篇
平均评分
年均产出4.0 篇/年
Ilia Kulikov
12
LLM Pretraining with Continuous Concepts
ICLR 2026Poster
14
J1: Incentivizing Thinking in LLM-as-a-Judge via Reinforcement Learning
ICLR 2026Poster
16
Hybrid Reinforcement: when reward is sparse, better to be dense
ICLR 2026Poster
二作12
OptimalThinkingBench: Evaluating Over and Underthinking in LLMs
ICLR 2026Poster
7
The Majority is not always right: RL training for solution aggregation
ICLR 2026Rejected
通讯13
Learning to Reason for Factuality
ICLR 2026Rejected
二作6
Adaptive Decoding via Latent Preference Optimization
ICLR 2026Rejected
二作6
Bridging Offline and Online Reinforcement Learning for LLMs
ICLR 2026Rejected
通讯5
Diverse Preference Optimization
ICLR 2026Rejected
通讯13
CoT-Self-Instruct: Building high-quality synthetic prompts data for reasoning and non-reasoning tasks
ICLR 2026Rejected