影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
96.46/100
前 0.2%
全站排名 #114
发表论文33 篇
平均评分
年均产出11.0 篇/年
Diyi Yang
研究方向
Computational Social Science · Responsible AI · Human Centered NLP
13
Relative Scaling Laws for LLMs
ICLR 2026Desk Rejected
通讯19
Real-Time Reasoning Agents in Evolving Environments
ICLR 2026Poster
23
Computer Agent Arena: Toward Human-Centric Evaluation and Analysis of Computer-Use Agents
ICLR 2026Poster
17
How Dark Patterns Manipulate Web Agents
ICLR 2026Poster
三作17
Optimas: Optimizing Compound AI Systems with Globally Aligned Local Rewards
ICLR 2026Poster
23
AutoLibra: Agent Metric Induction from Open-Ended Human Feedback
ICLR 2026Poster
通讯12
GEM: A Gym for Generalist LLMs
ICLR 2026Poster
13
The Ideation-Execution Gap: Execution Outcomes of LLM-Generated versus Human Research Ideas
ICLR 2026Poster
三作23
Searching for Privacy Risks in LLM Agents via Simulation
ICLR 2026Poster
二作25
Collaborative Gym: A Framework for Enabling and Evaluating Human-Agent Collaboration
ICLR 2026Poster
通讯15
AutoMetrics: Approximate Human Judgments with Automatically Generated Evaluators
ICLR 2026Poster
通讯12
ReplicationBench: Can AI Agents Replicate Astrophysics Research Papers?
ICLR 2026Rejected
22
OpenCUA: Open Foundations for Computer-Use Agents
NeurIPS 2025Spotlight
19
Blackbox Model Provenance via Palimpsestic Membership Inference
NeurIPS 2025Spotlight
15
The Unlearning Mirage: A Dynamic Framework for Evaluating LLM Unlearning
COLM 2025Poster
通讯26
Information Retrieval Induced Safety Degradation in AI Agents
NeurIPS 2025Poster
三作16
Internal Causal Mechanisms Robustly Predict Language Model Out-of-Distribution Behaviors
ICML 2025Poster
24
When Models Know More Than They Can Explain: Quantifying Knowledge Transfer in Human-AI Collaboration
NeurIPS 2025Poster
16
Aligning Language Models with Demonstrated Feedback
ICLR 2025Poster
通讯16
Can LLMs Generate Novel Research Ideas? A Large-Scale Human Study with 100+ NLP Researchers
ICLR 2025Poster
二作17
No Preference Left Behind: Group Distributional Preference Optimization
ICLR 2025Poster
14
SWE-bench Multimodal: Do AI Systems Generalize to Visual Software Domains?
ICLR 2025Poster
25
Distilling an End-to-End Voice Assistant Without Instruction Training Data
ICLR 2025Rejected
通讯5
Dynamic Skill Adaptation for Large Language Models
ICLR 2025Rejected
二作