影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
96.07/100
前 0.2%
全站排名 #137
发表论文34 篇
平均评分
年均产出11.3 篇/年
Wen Sun
研究方向
Imitation Learning · Reinforcement Learning
12
All Roads Lead to Likelihood: The Value of Reinforcement Learning in Fine-Tuning
ICLR 2026Poster
三作19
Prompt Curriculum Learning for Efficient LLM Post-Training
ICLR 2026Poster
三作16
Expressive Value Learning for Scalable Offline Reinforcement Learning
ICLR 2026Rejected
三作12
Value-as-Return: A Two-Stage Framework to Align on the Optimal Score Function
ICLR 2026Rejected
14
Controllable Diffusion via Optimal Classifier Guidance
ICLR 2026Rejected
通讯11
Test Time Scaling of Diffusion Model via Flow Matching Corrector
ICLR 2026Rejected
通讯14
Computationally Efficient RL under Linear Bellman Completeness for Deterministic Dynamics
ICLR 2025Oral
通讯9
$Q\sharp$: Provably Optimal Distributional RL for LLM Post-Training
ICML 2025Rejected
通讯20
Accelerating RL for LLM Reasoning with Optimal Advantage Regression
NeurIPS 2025Poster
22
$Q\sharp$: Provably Optimal Distributional RL for LLM Post-Training
NeurIPS 2025Poster
通讯19
Value-Guided Search for Efficient Chain-of-Thought Reasoning
NeurIPS 2025Poster
通讯22
Model-based RL as a Minimalist Approach to Horizon-Free and Second-Order Bounds
ICLR 2025Poster
通讯19
On Speeding Up Language Model Evaluation
ICLR 2025Poster
20
Avoiding exp(R) scaling in RLHF through Preference-based Exploration
NeurIPS 2025Poster
三作25
Diffusing States and Matching Scores: A New Framework for Imitation Learning
ICLR 2025Poster
通讯21
Scaling Offline RL via Efficient and Expressive Shortcut Models
NeurIPS 2025Poster
通讯16
Efficient Imitation under Misspecification
ICLR 2025Poster
三作11
Regressing the Relative Future: Efficient Policy Optimization for Multi-turn RLHF
ICLR 2025Poster
通讯22
Correcting the Mythos of KL-Regularization: Direct Alignment without Overoptimization via Chi-Squared Preference Optimization
ICLR 2025Spotlight
14
A Reductions Approach to Risk-Sensitive Reinforcement Learning with Optimized Certainty Equivalents
ICML 2025Poster
通讯17
Convergence Of Consistency Model With Multistep Sampling Under General Data Assumptions
ICLR 2025Rejected
通讯28
On Orchestrating Personalized LLMs
ICLR 2025Rejected
通讯10
Convergence of Consistency Model with Multistep Sampling under General Data Assumptions
ICML 2025Poster
通讯