影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
32.75/100
前 21.4%
全站排名 #13,760
发表论文6 篇
平均评分
年均产出3.0 篇/年
Eric Horvitz
研究方向
human-AI interaction and collaboration · augmenting human intelligence · decision-theoretic inference · utility-theoretic models · Markov decision processes · MDP · POMDP · cognitive science · human cognition · cognitive architectures · bounded rationality · bounded optimality · self-supervised training · semi-supervised training · causal inference · AI safety · reliability · robustness · adversarial machine learning · reward hacking · Bias · Fairness · Explanation of machine learning · transparency · Lifelong machine learning
21
Improving Instruction-Following in Language Models through Activation Steering
ICLR 2025Poster
25
Utility-Directed Conformal Prediction: A Decision-Aware Framework for Actionable Uncertainty Quantification
ICLR 2025Poster
12
MedFuzz: Exploring the Robustness of Large Language Models in Medical Question Answering
ICLR 2025Rejected
通讯