影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
82.58/100
前 1.2%
全站排名 #745
发表论文17 篇
平均评分
年均产出5.7 篇/年
Aldo Pacchiano
研究方向
Reinforcement Learning · Online Learning · Bandits · Optimization · Fairness
22
In-Context Learning for Pure Exploration
ICLR 2026Poster
三作15
Post-training Large Language Models for Diverse High-Quality Responses
ICLR 2026Poster
通讯6
Select the Right Agent: Data-Driven Online Model Selection in Reinforcement Learning
ICLR 2026Rejected
二作21
Learning with Coupled Uncertainty
ICLR 2026Rejected
三作5
Learning to Undo: Transfer Reinforcement Learning under State Space Transformations
ICLR 2026Withdrawn
二作12
Principled Fine-tuning of LLMs from User-Edits: A Medley of Preference, Supervision, and Reward
NeurIPS 2025Poster
二作12
Feasible Action Search for Bandit Linear Programs via Thompson Sampling
ICML 2025Poster
二作14
A Theoretical Framework for Partially-Observed Reward States in RLHF
ICLR 2025Poster
三作25
Second Order Bounds for Contextual Bandits with Function Approximation
ICLR 2025Poster
一作19
High Probability Contextual Bandits for Optimal Dosage Selection
ICLR 2025Rejected
二作14
Language Model Personalization via Reward Factorization
COLM 2025Poster
通讯44
ORSO: Accelerating Reward Design via Online Reward Selection and Policy Optimization
ICLR 2025Poster
三作9
Multiple-policy Evaluation via Density Estimation
ICML 2025Poster
二作15
Adaptive Exploration for Multi-Reward Multi-Policy Evaluation
ICML 2025Poster
二作15
Sample Efficient Multiple-policy Evaluation in Reinforcement Learning
ICLR 2025Rejected
二作