影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
37.89/100
前 16%
全站排名 #10,302
发表论文4 篇
平均评分
年均产出4.0 篇/年
11
Policy-labeled Preference Learning: Is Preference Enough for RLHF?
ICML 2025Spotlight
二作19
Pareto Optimal Risk-Agnostic Distributional Bandits with Heavy-Tail Rewards
NeurIPS 2025Poster
13
Bellman Unbiasedness: Toward Provably Efficient Distributional Reinforcement Learning with General Value Function Approximation
ICML 2025Poster
三作24
Bellman Unbiasedness: Toward Provably Efficient Distributional Reinforcement Learning with General Value Function Approximation
ICLR 2025Rejected