影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
78.41/100
前 1.5%
全站排名 #982
发表论文20 篇
平均评分
年均产出6.7 篇/年
21
AccidentBench: Benchmarking Multimodal Understanding and Reasoning in Vehicle Accidents and Beyond
ICLR 2026Rejected
16
Adversarial Déjà Vu: Jailbreak Dictionary Learning for Stronger Generalization to Unseen Attacks
ICLR 2026Poster
14
StyleBench: Evaluating thinking styles in Large Language Models
ICLR 2026Rejected
三作27
Don’t Trade Off Safety: Diffusion Regularization for Constrained Offline RL
NeurIPS 2025Poster
16
Reinforcement Learning with Backtracking Feedback
NeurIPS 2025Poster
42
Robust Gymnasium: A Unified Modular Benchmark for Robust Reinforcement Learning
ICLR 2025Poster
31
LLMs Can Plan Only If We Tell Them
ICLR 2025Poster
三作21
Probing Hidden Knowledge Holes in Unlearned LLMs
NeurIPS 2025Poster
12
LLMs Can Reason Faster Only If We Let Them
ICML 2025Poster
通讯33
A Black Swan Hypothesis: The Role of Human Irrationality in AI Safety
ICLR 2025Poster
通讯13
Just Enough Shifts: Mitigating Over-Refusal in Aligned Language Models with Targeted Representation Fine-Tuning
ICML 2025Poster
37
Data-Centric Human Preference Optimization with Rationales
ICLR 2025Rejected
二作22
LLM Spark: Critical Thinking Evaluation of Large Language Models
ICLR 2025Rejected
通讯18
Actions Speak Louder Than States: Going Beyond Bayesian Inference in In-Context Reinforcement Learning
ICLR 2025Rejected
通讯6
Rethinking the Uncertainty: A Critical Review and Analysis in the Era of Large Language Models
ICLR 2025Rejected