影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
97.29/100
前 0.1%
全站排名 #78
发表论文51 篇
平均评分
年均产出17.0 篇/年
14
Variational Reasoning for Language Models
ICLR 2026Poster
通讯13
Think in Parallel, Answer as One: Logit Averaging for Open-Ended Reasoning
ICLR 2026Poster
通讯35
Muon Outperforms Adam in Tail-End Associative Memory Learning
ICLR 2026Poster
12
Reinforcing Query-Level Meta-Agents
ICLR 2026Rejected
通讯24
Fostering Video Reasoning via Next-Event Prediction
ICLR 2026Poster
通讯21
Language Models Can Learn from Verbal Feedback Without Scalar Rewards
ICLR 2026Rejected
通讯22
Verl-Tool: Towards Holistic Agentic Reinforcement Learning with Tool Use
ICLR 2026Rejected
20
Nonparametric Data Attribution for Diffusion Models
ICLR 2026Rejected
18
Reinforcing General Reasoning Without Verifiers
ICLR 2026Poster
9
Anatomy of a Hybrid Mind: Deconstructing Hybrid Reasoning in Large Language Models
ICLR 2026Withdrawn
5
FreezeVLA: Action-Freezing Attacks against Vision-Language-Action Models
ICLR 2026Withdrawn
5
Imperceptible Jailbreaking against Large Language Models
ICLR 2026Withdrawn
通讯21
Orient Anything V2: Unifying Orientation and Rotation Understanding
NeurIPS 2025Spotlight
13
Understanding R1-Zero-Like Training: A Critical Perspective
COLM 2025Poster
40
Cheating Automatic LLM Benchmarks: Null Models Achieve High Win Rates
ICLR 2025Oral
二作16
When Attention Sink Emerges in Language Models: An Empirical View
ICLR 2025Spotlight
二作24
NoisyRollout: Reinforcing Visual Reasoning with Data Augmentation
NeurIPS 2025Poster
25
RegMix: Data Mixture as Regression for Language Model Pre-training
ICLR 2025Spotlight
15
Improving Your Model Ranking on Chatbot Arena by Vote Rigging
ICML 2025Poster
二作26
Adversarial Attacks against Closed-Source MLLMs via Feature Optimal Alignment
NeurIPS 2025Poster
16
Efficient Process Reward Model Training via Active Learning
COLM 2025Poster
25
Scaling up Masked Diffusion Models on Text
ICLR 2025Poster
23
Optimizing Anytime Reasoning via Budget Relative Policy Optimization
NeurIPS 2025Poster
三作22
SkyLadder: Better and Faster Pretraining via Context Window Scheduling
NeurIPS 2025Poster
24
Lifelong Safety Alignment for Language Models
NeurIPS 2025Poster
通讯25
Improved Techniques for Optimization-Based Jailbreaking on Large Language Models
ICLR 2025Poster
二作11
Weak-to-Strong Jailbreaking on Large Language Models
ICML 2025Poster
三作23
A Closer Look at Machine Unlearning for Large Language Models
ICLR 2025Poster
二作19
LLM-based Multi-Agents System Attack via Continuous Optimization with Discrete Efficient Search
COLM 2025Poster
三作28
Bootstrapping Language Models with DPO Implicit Rewards
ICLR 2025Poster
27
Improving Long-Text Alignment for Text-to-Image Diffusion Models
ICLR 2025Poster
三作13
Orient Anything: Learning Robust Object Orientation Estimation from Rendering 3D Models
ICML 2025Poster
三作41
Weak-to-Strong Jailbreaking on Large Language Models
ICLR 2025Rejected
三作26
SimLayerKV: A Simple Framework for Layer-Level KV Cache Reduction
ICLR 2025Rejected
15
Unnatural Languages Are Not Bugs but Features for LLMs
ICML 2025Poster
10
BanditSpec: Adaptive Speculative Decoding via Bandit Algorithms
ICML 2025Poster
6
Test-Time Backdoor Attacks on Multimodal Large Language Models
ICLR 2025Withdrawn
二作5
Meta-Unlearning on Diffusion Models: Preventing Relearning Unlearned Concepts
ICLR 2025Withdrawn
二作17
Denial-of-Service Poisoning Attacks against Large Language Models
ICLR 2025Withdrawn
二作