影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
88.5/100
前 0.7%
全站排名 #436
发表论文27 篇
平均评分
年均产出9.0 篇/年
Yali Du
研究方向
LLM agents · Reinforcement Learning · Multi-agent learnining · Robust machine learning · Machine learning
24
Is Pure Exploitation Sufficient in Exogenous MDPs with Linear Function Approximation?
ICLR 2026Poster
通讯22
SocialJax: An Evaluation Suite for Multi-agent Reinforcement Learning in Sequential Social Dilemmas
ICLR 2026Poster
通讯21
BRIDGE: Bi-level Reinforcement Learning for Dynamic Group Structure in Coalition Formation Games
ICLR 2026Poster
通讯15
Learning Instruction-Following Policies through Open-Ended Instruction Relabeling with Large Language Models
ICLR 2026Rejected
三作11
Intrinsic Memory Agents: Heterogeneous Multi-Agent LLM Systems through Structured Contextual Memory
ICLR 2026Rejected
28
MEAL: A Benchmark for Continual Multi-Agent Reinforcement Learning
ICLR 2026Rejected
14
Distill Not Only Data but Also Rewards: Can Smaller Language Models Surpass Larger Ones?
ICLR 2026Rejected
11
Policy Transfer for Improved Sample Efficiency in Goal-Conditioned Reinforcement Learning
ICLR 2026Withdrawn
通讯19
Causality Meets Locality: Provably Generalizable and Scalable Policy Learning for Networked Systems
NeurIPS 2025Spotlight
通讯22
Social World Model-Augmented Mechanism Design Policy Learning
NeurIPS 2025Poster
17
Abstract Counterfactuals for Language Model Agents
NeurIPS 2025Poster
三作29
Self-Verifying Reflection Helps Transformers with CoT Reasoning
NeurIPS 2025Poster
11
M³HF: Multi-agent Reinforcement Learning from Multi-phase Human Feedback of Mixed Quality
ICML 2025Poster
通讯26
On the Optimization Landscape of Low Rank Adaptation Methods for Large Language Models
ICLR 2025Poster
二作20
RuAG: Learned-rule-augmented Generation for Large Language Models
ICLR 2025Poster
15
GRU: Mitigating the Trade-off between Unlearning and Retention for LLMs
ICML 2025Poster
22
VLP: Vision-Language Preference Learning for Embodied Manipulation
ICLR 2025Rejected
9
Resolving Complex Social Dilemmas by Aligning Preferences with Counterfactual Regret
ICLR 2025Rejected
通讯