影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
96.2/100
前 0.2%
全站排名 #130
发表论文52 篇
平均评分
年均产出17.3 篇/年
14
Variational Reasoning for Language Models
ICLR 2026Poster
13
Think in Parallel, Answer as One: Logit Averaging for Open-Ended Reasoning
ICLR 2026Poster
二作35
Muon Outperforms Adam in Tail-End Associative Memory Learning
ICLR 2026Poster
12
Reinforcing Query-Level Meta-Agents
ICLR 2026Rejected
24
Fostering Video Reasoning via Next-Event Prediction
ICLR 2026Poster
21
Language Models Can Learn from Verbal Feedback Without Scalar Rewards
ICLR 2026Rejected
22
Verl-Tool: Towards Holistic Agentic Reinforcement Learning with Tool Use
ICLR 2026Rejected
20
Nonparametric Data Attribution for Diffusion Models
ICLR 2026Rejected
二作5
Sample-Efficient Alignment for LLMs
ICLR 2026Rejected
三作18
Reinforcing General Reasoning Without Verifiers
ICLR 2026Poster
通讯9
Anatomy of a Hybrid Mind: Deconstructing Hybrid Reasoning in Large Language Models
ICLR 2026Withdrawn
5
Imperceptible Jailbreaking against Large Language Models
ICLR 2026Withdrawn
三作5
FreezeVLA: Action-Freezing Attacks against Vision-Language-Action Models
ICLR 2026Withdrawn
21
Orient Anything V2: Unifying Orientation and Rotation Understanding
NeurIPS 2025Spotlight
13
Understanding R1-Zero-Like Training: A Critical Perspective
COLM 2025Poster
40
Cheating Automatic LLM Benchmarks: Null Models Achieve High Win Rates
ICLR 2025Oral
三作16
When Attention Sink Emerges in Language Models: An Empirical View
ICLR 2025Spotlight
三作24
NoisyRollout: Reinforcing Visual Reasoning with Data Augmentation
NeurIPS 2025Poster
11
Continual Reinforcement Learning by Planning with Online World Models
ICML 2025Spotlight
三作15
Improving Your Model Ranking on Chatbot Arena by Vote Rigging
ICML 2025Poster
三作26
Adversarial Attacks against Closed-Source MLLMs via Feature Optimal Alignment
NeurIPS 2025Poster
25
Scaling up Masked Diffusion Models on Text
ICLR 2025Poster
三作23
Optimizing Anytime Reasoning via Budget Relative Policy Optimization
NeurIPS 2025Poster
24
Lifelong Safety Alignment for Language Models
NeurIPS 2025Poster
25
Improved Techniques for Optimization-Based Jailbreaking on Large Language Models
ICLR 2025Poster
三作11
Weak-to-Strong Jailbreaking on Large Language Models
ICML 2025Poster
28
Bootstrapping Language Models with DPO Implicit Rewards
ICLR 2025Poster
三作23
A Closer Look at Machine Unlearning for Large Language Models
ICLR 2025Poster
三作19
LLM-based Multi-Agents System Attack via Continuous Optimization with Discrete Efficient Search
COLM 2025Poster
16
Sample-Efficient Alignment for LLMs
NeurIPS 2025Rejected
三作27
Improving Long-Text Alignment for Text-to-Image Diffusion Models
ICLR 2025Poster
二作13
Orient Anything: Learning Robust Object Orientation Estimation from Rendering 3D Models
ICML 2025Poster
41
Weak-to-Strong Jailbreaking on Large Language Models
ICLR 2025Rejected
26
SimLayerKV: A Simple Framework for Layer-Level KV Cache Reduction
ICLR 2025Rejected
三作19
Sample Efficient Alignment for LLMs
ICLR 2025Rejected
三作10
BanditSpec: Adaptive Speculative Decoding via Bandit Algorithms
ICML 2025Poster
6
Test-Time Backdoor Attacks on Multimodal Large Language Models
ICLR 2025Withdrawn
三作5
Meta-Unlearning on Diffusion Models: Preventing Relearning Unlearned Concepts
ICLR 2025Withdrawn
三作17
Denial-of-Service Poisoning Attacks against Large Language Models
ICLR 2025Withdrawn
三作