影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
93.07/100
前 0.4%
全站排名 #248
发表论文49 篇
平均评分
年均产出16.3 篇/年
William Yang Wang
研究方向
Vision and Language · Computational Social Science · Information Extraction
12
LogicReward: Incentivizing LLM Reasoning via Step-Wise Logical Supervision
ICLR 2026Poster
19
DevOps-Gym: Benchmarking AI Agents in Software DevOps Cycle
ICLR 2026Poster
25
Dynamic Speculative Agent Planning
ICLR 2026Poster
15
SOPBench: Evaluating Language Agents at Following Standard Operating Procedures and Constraints
ICLR 2026Rejected
17
Adversarial Training for Process Reward Models
ICLR 2026Desk Rejected
三作13
Do Larger Language Models Generalize Better? A Scaling Law for Implicit Reasoning at Pretraining Time
ICLR 2026Rejected
15
Self-Resource Allocation in Multi-Agent LLM Systems
ICLR 2026Rejected
通讯16
PromptArmor: An Essential Baseline for Prompt Injection Defenses
ICLR 2026Rejected
14
Cost-effective Agent Test-time Scaling via Budget-Aware Thinking
ICLR 2026Rejected
5
HexMachina: Self-Evolving Multi-Agent System for Continual Learning of Catan
ICLR 2026Withdrawn
通讯25
MuSLR: Multimodal Symbolic Logical Reasoning
NeurIPS 2025Poster
11
Weak-to-Strong Jailbreaking on Large Language Models
ICML 2025Poster
通讯18
ThoughtTerminator: Benchmarking, Calibrating, and Mitigating Overthinking in Reasoning Models
COLM 2025Poster
通讯28
Speculative Knowledge Distillation: Bridging the Teacher-Student Gap Through Interleaved Sampling
ICLR 2025Poster
43
T2V-Turbo-v2: Enhancing Video Model Post-Training through Data, Reward, and Conditional Guidance Design
ICLR 2025Poster
通讯39
MMWorld: Towards Multi-discipline Multi-faceted World Model Evaluation in Videos
ICLR 2025Poster
38
MMSci: A Dataset for Graduate-Level Multi-Discipline Multimodal Scientific Understanding
ICLR 2025Rejected
通讯26
COrAL: Order-Agnostic Language Modeling for Efficient Iterative Refinement
ICLR 2025Rejected
通讯19
MLGym: A New Framework and Benchmark for Advancing AI Research Agents
COLM 2025Poster
52
Gödel Agent: A Self-Referential Framework Helps for Recursively Self-Improvement
ICLR 2025Rejected
通讯41
Weak-to-Strong Jailbreaking on Large Language Models
ICLR 2025Rejected
通讯20
Discovering Factor Level Preferences to Improve Human-Model Alignment
ICLR 2025Rejected
31
Understanding the Interplay between Parametric and Contextual Knowledge for Large Language Models
ICLR 2025Rejected
通讯22
Generalization v.s. Memorization: Tracing Language Models’ Capabilities Back to Pretraining Data
ICLR 2025Poster
通讯5
VSP: Assessing the dual challenges of perception and reasoning in spatial planning tasks for MLLMs
ICLR 2025Withdrawn
6
TC-Bench: Benchmarking Temporal Compositionality in Conditional Video Generation
ICLR 2025Rejected
通讯7
Can Editing LLMs Inject Harm?
ICLR 2025Rejected
11
MELON: Provable Defense Against Indirect Prompt Injection Attacks in AI Agents
ICML 2025Poster
通讯5
Compact Multimodal Context Represenations Using Visual Tokens
ICLR 2025Rejected
通讯14
Pixelated Instructions: Can Multimodal Large Language Models Follow Printed Instructions in Images?
ICLR 2025Rejected
28
SWE-Search: Enhancing Software Agents with Monte Carlo Tree Search and Iterative Refinement
ICLR 2025Poster
通讯17
Detecting Training Data of Large Language Models via Expectation Maximization
ICLR 2025Rejected
通讯4
DebUnc: Improving Large Language Model Agent Communication Via Uncertainty Metrics
ICLR 2025Withdrawn
三作