影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
26.59/100
前 28.8%
全站排名 #18,531
发表论文6 篇
平均评分
年均产出2.0 篇/年
Ming Shi
研究方向
Reinforcement learning · Bandit learning · Online Convex Optimization
15
Provably Efficient RL for Linear MDPs under Instantaneous Safety Constraints in Non-Convex Feature Spaces
ICML 2025Poster
三作31
Provably Efficient Multi-Objective Bandit Algorithms under Preference-Centric Customization
ICLR 2025Rejected
二作29
RLHF with Inconsistent Multi-Agent Feedback Under General Function Approximation: A Theoretical Perspective
ICLR 2025Rejected
一作33
Provably Efficient Linear Bandits with Instantaneous Constraints in Non-Convex Feature Spaces
ICLR 2025Rejected
二作