影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
68.5/100
前 2.9%
全站排名 #1,859
发表论文17 篇
平均评分
年均产出5.7 篇/年
Zhiheng Xi
研究方向
Large Language Models · LLM Reasoning · LLM Agent · LLM RL
22
AgentGym-RL: An Open-Source Framework to Train LLM Agents for Long-Horizon Decision Making via Multi-Turn RL
ICLR 2026Oral
一作22
Game-RL: Synthesizing Multimodal Verifiable Game Data to Boost VLMs' General Reasoning
ICLR 2026Poster
28
BAPO: Stabilizing Off-Policy Reinforcement Learning for LLMs via Balanced Policy Optimization with Adaptive Clipping
ICLR 2026Poster
一作19
Critique-RL: Training Language Models For Critiquing Through Two-Stage Reinforcement Learning
ICLR 2026Poster
一作28
Better know nothing than half-know anything: A Precise and Efficient Dataset for Scientific Reasoning in Language Models
ICLR 2026Rejected
三作17
Why Reinforcement Fine-Tuning Enables MLLMs Preserve Prior Knowledge Better: A Data Perspective
ICLR 2026Poster
20
Two Minds Better Than One: Collaborative Reward Modeling for LLM Alignment
ICLR 2026Rejected
13
From Scores to Preferences: Redefining MOS Benchmarking for Speech Quality Reward Modeling
ICLR 2026Withdrawn
24
Pre-Trained Policy Discriminators are General Reward Models
NeurIPS 2025Poster
14
Have the VLMs Lost Confidence? A Study of Sycophancy in VLMs
ICLR 2025Poster
26
Distill Visual Chart Reasoning Ability from LLMs to MLLMs
ICLR 2025Rejected
二作22
RMB: Comprehensively benchmarking reward models in LLM alignment
ICLR 2025Poster
44
AgentGym: Evaluating and Evolving Large Language Model-based Agents across Diverse Envronments
ICLR 2025Rejected
一作15
Progressive LLM Alignments Using Two-Player Games
ICLR 2025Rejected