影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
46.62/100
前 10.1%
全站排名 #6,474
发表论文13 篇
平均评分
年均产出4.3 篇/年
14
J1: Incentivizing Thinking in LLM-as-a-Judge via Reinforcement Learning
ICLR 2026Poster
二作16
Hybrid Reinforcement: when reward is sparse, better to be dense
ICLR 2026Poster
15
Jointly Reinforcing Diversity and Quality in Language Model Generations
ICLR 2026Rejected
通讯6
Bridging Offline and Online Reinforcement Learning for LLMs
ICLR 2026Rejected
13
CoT-Self-Instruct: Building high-quality synthetic prompts data for reasoning and non-reasoning tasks
ICLR 2026Rejected
三作5
ASTRO: Teaching Language Models to Reason by Reflecting and Backtracking In-Context
ICLR 2026Withdrawn
通讯18
Multi-Token Attention
COLM 2025Poster
二作19
Contextual Position Encoding: Learning to Count What’s Important
ICLR 2025Rejected
二作15
Self-Taught Evaluators
ICLR 2025Rejected
一作11
Learning to Plan & Reason for Evaluation with Thinking-LLM-as-a-Judge
ICML 2025Poster
通讯9
Calibrate to Discriminate:Improve In-Context Learning with Label-Free Comparative Inference
ICLR 2025Rejected
二作