影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
60.44/100
前 4.7%
全站排名 #3,047
发表论文25 篇
平均评分
年均产出12.5 篇/年
Amrit Singh Bedi
研究方向
AI Alignment · Language models · Reinforcement Learning · policy gradient algorithms
20
Direct Preference Optimization for Primitive-Enabled Hierarchical RL: A Bilevel Approach
ICLR 2026Poster
通讯14
Repair Aware Forgetting: An Iterative Approach to Unlearning in T2I Diffusion Models
ICLR 2026Desk Rejected
27
TEST-TIME SCALING IN DIFFUSION LLMS VIA HIDDEN SEMI-AUTOREGRESSIVE EXPERTS
ICLR 2026Poster
通讯26
Multi-Level Multi-Turn RL Outperforms GRPO: Reasoning with Textual Feedback
ICLR 2026Rejected
通讯14
Cut the Overcredit: Precision First Process Rewards for Reasoning LLMs
ICLR 2026Rejected
4
Advancing Regulation in Artificial Intelligence: An Auction-Based Approach
ICLR 2026Withdrawn
29
TRAM: Test-time Risk Adaptation with Mixture of Agents
ICLR 2026Rejected
二作13
Improved Sample Complexity Bounds For Diffusion Model Training Without Empirical Risk Minimizer Access
ICLR 2026Withdrawn
25
Mitigating Reward Hacking in Inference-Time Alignment of T2I Diffusion Models via Distributional Regularization
ICLR 2026Rejected
13
A Principled Approach to Chain-of-Thought Monitorability in Reasoning Models
ICLR 2026Withdrawn
三作4
SafeThink: A Key to Safety in Multi-Modal Large Reasoning Models
ICLR 2026Withdrawn
通讯21
On the Global Optimality of Policy Gradient Methods in General Utility Reinforcement Learning
NeurIPS 2025Poster
通讯23
Does Thinking More Always Help? Mirage of Test-Time Scaling in Reasoning Models
NeurIPS 2025Poster
通讯24
On the Sample Complexity Bounds of Bilevel Reinforcement Learning
NeurIPS 2025Poster
三作11
Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time
ICML 2025Poster
通讯22
SAIL: Self-improving Efficient Online Alignment of Large Language Models
ICLR 2025Rejected
8
Hierarchical Preference Optimization: Learning to achieve goals via feasible subgoals prediction
ICLR 2025Withdrawn
通讯22
Towards Realistic Mechanisms That Incentivize Federated Participation and Contribution
ICLR 2025Rejected
二作8
Aligning Large Language Models With Preference Privacy
ICLR 2025Rejected
30
Auction-Based Regulation for Artificial Intelligence
ICLR 2025Rejected
30
LIAR: Leveraging Inverse Alignment to Jailbreak LLMs in Seconds
ICLR 2025Rejected
26
On the Sample Complexity of a Policy Gradient Algorithm with Occupancy Approximation for General Utility Reinforcement Learning
ICLR 2025Rejected
通讯5
DIPPER: Direct Preference Optimization for Primitive-Enabled Hierarchical Reinforcement Learning
ICLR 2025Withdrawn
通讯5
On the Global Convergence of RLHF Based Alignment With Neural Parametrization
ICLR 2025Withdrawn
二作5
AIME: AI System Optimization via Multiple LLM Evaluators
ICLR 2025Withdrawn