影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
46.62/100
前 10.1%
全站排名 #6,475
发表论文8 篇
平均评分
年均产出2.7 篇/年
Katherine Metcalf
研究方向
human in the loop learning · human robot interaction · reinforcement learning from human feedback · RLHF · preference learning · alignment · PbRL · preference-based reinforcement learning
19
Investigating Intersectional Bias in Large Language Models using Confidence Disparities in Coreference Resolution
COLM 2025Poster
8
Aligning LLMs by Predicting Preferences from User Writing Samples
ICML 2025Poster
通讯11
Is Your Model Fairly Certain? Uncertainty-Aware Fairness Evaluation for LLMs
ICML 2025Poster
30
PREDICT: Preference Reasoning by Evaluating Decomposed preferences Inferred from Candidate Trajectories
ICLR 2025Rejected
通讯5
Uncovering Intersectional Stereotypes in Humans and Large Language Models
ICLR 2025Withdrawn