影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
75.58/100
前 1.8%
全站排名 #1,166
发表论文15 篇
平均评分
年均产出5.0 篇/年
Zhiwei Zhang
研究方向
Reinforcement Learning · Trustworthy AI
29
Multi-Head Low-Rank Attention
ICLR 2026Poster
三作23
Multiplayer Nash Preference Optimization
ICLR 2026Oral
29
Unlocking the Power of Multi-Agent LLM for Reasoning: From Lazy Agents to Deliberation
ICLR 2026Poster
一作20
Agent-REINFORCE: Searching Compute-Optimal Multi-LLM Collaboration Graph for Test-Time Scaling
ICLR 2026Rejected
27
How Far Are LLMs from Professional Poker Players? Revisiting Game-Theoretic Reasoning with Agentic Tool Use
ICLR 2026Poster
30
Bradley-Terry and Multi-Objective Reward Modeling Are Complementary
ICLR 2026Poster
一作14
Turn-Level Trajectory Optimization for Robust Multi-Turn LLM Reasoning
ICLR 2026Withdrawn
通讯24
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
NeurIPS 2025Spotlight
三作21
Robustness Inspired Graph Backdoor Defense
ICLR 2025Oral
一作22
AgentTTS: Large Language Model Agent for Test-time Compute-optimal Scaling Strategy in Complex Tasks
NeurIPS 2025Poster
29
Catastrophic Failure of LLM Unlearning via Quantization
ICLR 2025Poster
一作33
Rule-Based Rating and Selection of LLM Training Data
ICLR 2025Rejected
三作7
RuleAdapter: Dynamic Rules for training Safety Reward Models in RLHF
ICML 2025Poster
三作