影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
31.02/100
前 23.3%
全站排名 #15,011
发表论文6 篇
平均评分
年均产出2.0 篇/年
Siliang Zeng
研究方向
Reinforcement Learning · Inverse Reinforcement Learning · Large Language Model
18
Joint Reward and Policy Learning with Demonstrations and Human Feedback Improves Alignment
ICLR 2025Spotlight
二作16
From Demonstrations to Rewards: Alignment Without Explicit Human Preferences
ICLR 2025Rejected
一作6
Policy optimization can be memory-efficient: LLM Alignment Through Successive Policy Re-weighting (SPR)
ICLR 2025Rejected
二作