影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
81.89/100
前 1.2%
全站排名 #781
发表论文14 篇
平均评分
年均产出7.0 篇/年
15
ARM-FM: Automated Reward Machines via Foundation Models for Compositional Reinforcement Learning
ICLR 2026Poster
三作20
Simplicial Embeddings Improve Sample Efficiency in Actor–Critic Agents
ICLR 2026Poster
通讯18
A Comedy of Estimators: On KL Regularization in RL Training of LLMs
ICLR 2026Rejected
16
Asymmetric Proximal Policy Optimization: mini-critics boost LLM reasoning
ICLR 2026Poster
18
Align and Filter: Improving Performance in Asynchronous On-Policy RL
ICLR 2026Rejected
18
Stable Gradients for Stable Learning at Scale in Deep Reinforcement Learning
NeurIPS 2025Spotlight
通讯15
Discovering Symbolic Cognitive Models from Human and Animal Behavior
ICML 2025Spotlight
一作23
Don't flatten, tokenize! Unlocking the key to SoftMoE's efficacy in deep RL
ICLR 2025Spotlight
通讯18
Mind the GAP! The Challenges of Scale in Pixel-based Deep Reinforcement Learning
NeurIPS 2025Poster
二作26
Measure gradients, not activations! Enhancing neuronal activity in deep reinforcement learning
NeurIPS 2025Poster
22
Studying the Interplay Between the Actor and Critic Representations in Reinforcement Learning
ICLR 2025Poster
三作17
The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning
ICML 2025Poster
三作15
Mitigating Plasticity Loss in Continual Reinforcement Learning by Reducing Churn
ICML 2025Poster
三作11
The Impact of On-Policy Parallelized Data Collection on Deep Reinforcement Learning Networks
ICML 2025Poster
通讯