影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
71.52/100
前 2.4%
全站排名 #1,528
发表论文28 篇
平均评分
年均产出9.3 篇/年
Dongmei Zhang
研究方向
Software Analytics · Computer Vision
21
Pretrain Value, Not Reward: Decoupled Value Policy Optimization
ICLR 2026Poster
23
DoVer: Intervention-Driven Auto Debugging for LLM Multi-Agent Systems
ICLR 2026Poster
通讯19
RePrompt: Reasoning-Augmented Reprompting for Text-to-Image Generation via Reinforcement Learning
ICLR 2026Poster
30
Text2Grad: Reinforcement Learning from Natural Language Feedback
ICLR 2026Poster
通讯14
SuperRL: Reinforcement Learning with Supervision to Boost Language Model Reasoning
ICLR 2026Rejected
通讯31
Test Time Training for Supervised Causal Learning
ICLR 2026Rejected
通讯16
WarriorMath: Empowering Mathematical Reasoning for Large Language Models via Expert Battles
ICLR 2026Rejected
通讯14
Towards Reliable Transferability of Targeted Adversarial Attacks against Model Discrepancy
ICLR 2026Rejected
9
Fortune: Formula-Driven Reinforcement Learning for Symbolic Table Reasoning in Language Models
ICLR 2026Withdrawn
通讯5
Camouflage Patching: Effective Jailbreak Attacks on Single- and Multimodal LLMs
ICLR 2026Withdrawn
14
Memory-Augmented Personalized Retrieval for Long-Context Egocentric Video
ICLR 2026Withdrawn
22
GUI-360° : A Comprehensive Dataset and Benchmark for Computer-Using Agents
ICLR 2026Rejected
通讯14
Learning to Refine: Self-Refinement of Parallel Reasoning in LLMs
ICLR 2026Withdrawn
通讯46
G-KV: Decoding-Time KV Cache Eviction with Global Attention
ICLR 2026Withdrawn
5
Duet: Joint Exploration of User–Item Profiles
ICLR 2026Withdrawn
通讯6
VEM: Environment-Free Exploration for Training GUI Agent with Value Environment Model
ICLR 2026Withdrawn
5
Bingo: Boosting Efficient Reasoning of LLMs via Dynamic and Significance-based Reinforcement Learning
ICLR 2026Withdrawn
通讯14
Distill Not Only Data but Also Rewards: Can Smaller Language Models Surpass Larger Ones?
ICLR 2026Rejected
30
WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct
ICLR 2025Oral
通讯19
Consensus-Robust Transfer Attacks via Parameter and Representation Perturbations
NeurIPS 2025Poster
25
GUI-Actor: Coordinate-Free Visual Grounding for GUI Agents
NeurIPS 2025Poster
40
OpenRCA: Can Large Language Models Locate the Root Cause of Software Failures?
ICLR 2025Poster
20
RuAG: Learned-rule-augmented Generation for Large Language Models
ICLR 2025Poster
37
SELF-EVOLVED REWARD LEARNING FOR LLMS
ICLR 2025Poster
29
Thread: A Logic-Based Data Organization Paradigm for How-To Question Answering with Retrieval Augmented Generation
ICLR 2025Rejected
29
Test-Time Learning of Causal Structure from Interventional Data
ICLR 2025Rejected
通讯6
Zoomer: Enhancing MLLM Performance with Adaptive Image Focus Optimization
ICLR 2025Withdrawn