影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
34.47/100
前 19.5%
全站排名 #12,547
发表论文6 篇
平均评分
年均产出3.0 篇/年
Tomasz Korbak
研究方向
language models · reinforcement learning from human feedback · AI safety · generative models · variational inference · probabilistic programming · emergent communication · compositionality
23
Towards Understanding Sycophancy in Language Models
ICLR 2024Poster
三作25
Is Model Collapse Inevitable? Breaking the Curse of Recursion by Accumulating Real and Synthetic Data
COLM 2024Poster
19
The Reversal Curse: LLMs trained on “A is B” fail to learn “B is A”
ICLR 2024Poster
20
Many-shot Jailbreaking
NeurIPS 2024Poster
18
Compositional Preference Models for Aligning LMs
ICLR 2024Poster
二作