影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
95.79/100
前 0.2%
全站排名 #149
发表论文36 篇
平均评分
年均产出12.0 篇/年
Amy Zhang
研究方向
Deep Reinforcement Learning · Planning · Representation Learning
11
Regularized Latent Dynamics Prediction is a Strong Baseline For Behavioral Foundation Models
ICLR 2026Poster
19
Reinforcement Learning via Value Gradient Flow
ICLR 2026Poster
通讯15
Hierarchical Entity-centric Reinforcement Learning with Factored Subgoal Diffusion
ICLR 2026Poster
18
Multi-agent Coordination via Flow Matching
ICLR 2026Poster
三作18
Reevaluating Policy Gradient Methods for Imperfect-Information Games
ICLR 2026Poster
23
Self-Refining Vision Language Model for Robotic Failure Detection and Reasoning
ICLR 2026Poster
23
A Unifying Perspective on Unsupervised Reinforcement Learning Algorithms
ICLR 2026Rejected
通讯24
Learning to Interact in World Latent for Team Coordination
ICLR 2026Rejected
29
TRAM: Test-time Risk Adaptation with Mixture of Agents
ICLR 2026Rejected
10
CARE-RFT: Confidence-Anchored Reinforcement Finetuning for Reliable Reasoning in Large Language Models
ICLR 2026Rejected
通讯12
Exploration for Deployment-Efficient Reinforcement Learning Agents
ICLR 2026Rejected
12
FLAM: Scaling Latent Action World Models with Factorization
ICLR 2026Rejected
27
MaestroMotif: Skill Design from Artificial Intelligence Feedback
ICLR 2025Oral
17
Towards General-Purpose Model-Free Reinforcement Learning
ICLR 2025Spotlight
三作16
Information-Theoretic Reward Decomposition for Generalizable RLHF
NeurIPS 2025Poster
三作24
RLZero: Direct Policy Inference from Language Without In-Domain Supervision
NeurIPS 2025Poster
19
ExPO: Unlocking Hard Reasoning with Self-Explanation-Guided Reinforcement Learning
NeurIPS 2025Poster
三作25
Uni-RL: Unifying Online and Offline RL via Implicit Value Regularization
NeurIPS 2025Poster
通讯29
Proto Successor Measure: Representing the space of all possible solutions of Reinforcement Learning
ICLR 2025Rejected
通讯26
Null Counterfactual Factor Interactions for Goal-Conditioned Reinforcement Learning
ICLR 2025Poster
19
Learning a Fast Mixing Exogenous Block MDP using a Single Trajectory
ICLR 2025Poster
三作31
EC-Diffuser: Multi-Object Manipulation via Entity-Centric Behavior Generation
ICLR 2025Poster
通讯13
Proto Successor Measure: Representing the Behavior Space of an RL Agent
ICML 2025Poster
通讯23
An Optimal Discriminator Weighted Imitation Perspective for Reinforcement Learning
ICLR 2025Poster
通讯23
Online Intrinsic Rewards for Decision Making Agents from Large Language Model Feedback
ICLR 2025Rejected
三作5
Augmented Conditioning is Enough for Effective Training Image Generation
ICLR 2025Withdrawn
二作