影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
75.19/100
前 1.9%
全站排名 #1,196
发表论文20 篇
平均评分
年均产出6.7 篇/年
Zhaoran Wang
18
Beyond Markovian: Reflective Exploration via Bayes-Adaptive RL for LLM Reasoning
ICLR 2026Poster
13
Learning to Reason as Action Abstractions with Scalable Mid-Training RL
ICLR 2026Poster
19
Local Linear Attention: An Optimal Interpolation of Linear and Softmax Attention For Test-Time Regression
ICLR 2026Poster
通讯27
Are Transformers Able to Reason by Connecting Separated Knowledge in Training Data?
ICLR 2025Poster
二作13
Reward-Augmented Data Enhances Direct Preference Alignment of LLMs
ICML 2025Poster
通讯8
BRiTE: Bootstrapping Reinforced Thinking Process to Enhance Language Model Reasoning
ICML 2025Poster
通讯34
Reward-Augmented Data Enhances Direct Preference Alignment of LLMs
ICLR 2025Rejected
通讯11
The Sample Complexity of Online Strategic Decision Making with Information Asymmetry and Knowledge Transportability
ICML 2025Poster
15
Progressive LLM Alignments Using Two-Player Games
ICLR 2025Rejected
18
Provably Efficient and Practical Self-Play for Better LLM Alignment
ICLR 2025Rejected
通讯5
How Can LLM Guide RL? A Value-Based Approach
ICLR 2025Withdrawn
通讯9
An Instrumental Value for Data Production and its Application to Data Pricing
ICML 2025Poster
三作6
Human-Instruction-Free LLM Self-Alignment with Limited Samples
ICLR 2025Rejected
-1
Hindsight Planner: A Closed-loop few-shot planner for Embodied Instruction Following
ICLR 2025Withdrawn
通讯-1
Self-Exploring Language Models: Active Preference Elicitation for Online Alignment
ICLR 2025Withdrawn
通讯