影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
98.67/100
前 0.1%
全站排名 #29
发表论文49 篇
平均评分
年均产出16.3 篇/年
Min Lin
研究方向
XLA Compiler · Quantum Chemistry · Information Theory · Bayesian Deep Learning · Generative Models · Continual Learning · Deep Learning Systems · Dynamical Systems · GAN · Convolutional Neural Networks
14
Variational Reasoning for Language Models
ICLR 2026Poster
18
Revisiting Parameter Server in LLM Post-Training
ICLR 2026Poster
16
SPIRAL: Self-Play on Zero-Sum Games Incentivizes Reasoning via Multi-Agent Multi-Turn Reinforcement Learning
ICLR 2026Poster
12
GEM: A Gym for Generalist LLMs
ICLR 2026Poster
通讯12
Reinforcing Query-Level Meta-Agents
ICLR 2026Rejected
21
Language Models Can Learn from Verbal Feedback Without Scalar Rewards
ICLR 2026Rejected
16
Flow-Distorted Plane Waves
ICLR 2026Rejected
20
Nonparametric Data Attribution for Diffusion Models
ICLR 2026Rejected
通讯5
Sample-Efficient Alignment for LLMs
ICLR 2026Rejected
通讯18
Reinforcing General Reasoning Without Verifiers
ICLR 2026Poster
13
Understanding R1-Zero-Like Training: A Critical Perspective
COLM 2025Poster
通讯40
Cheating Automatic LLM Benchmarks: Null Models Achieve High Win Rates
ICLR 2025Oral
通讯16
When Attention Sink Emerges in Language Models: An Empirical View
ICLR 2025Spotlight
通讯25
RegMix: Data Mixture as Regression for Language Model Pre-training
ICLR 2025Spotlight
通讯11
Continual Reinforcement Learning by Planning with Online World Models
ICML 2025Spotlight
通讯15
Improving Your Model Ranking on Chatbot Arena by Vote Rigging
ICML 2025Poster
通讯25
Scaling up Masked Diffusion Models on Text
ICLR 2025Poster
24
Lifelong Safety Alignment for Language Models
NeurIPS 2025Poster
23
Optimizing Anytime Reasoning via Budget Relative Policy Optimization
NeurIPS 2025Poster
通讯25
Improved Techniques for Optimization-Based Jailbreaking on Large Language Models
ICLR 2025Poster
通讯9
PipeOffload: Improving Scalability of Pipeline Parallelism with Memory Optimization
ICML 2025Poster
19
LLM-based Multi-Agents System Attack via Continuous Optimization with Discrete Efficient Search
COLM 2025Poster
23
A Closer Look at Machine Unlearning for Large Language Models
ICLR 2025Poster
通讯28
Bootstrapping Language Models with DPO Implicit Rewards
ICLR 2025Poster
通讯16
Sample-Efficient Alignment for LLMs
NeurIPS 2025Rejected
通讯26
SimLayerKV: A Simple Framework for Layer-Level KV Cache Reduction
ICLR 2025Rejected
通讯19
Sample Efficient Alignment for LLMs
ICLR 2025Rejected
通讯6
Test-Time Backdoor Attacks on Multimodal Large Language Models
ICLR 2025Withdrawn
通讯17
Denial-of-Service Poisoning Attacks against Large Language Models
ICLR 2025Withdrawn
通讯5
Meta-Unlearning on Diffusion Models: Preventing Relearning Unlearned Concepts
ICLR 2025Withdrawn
通讯