影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
97.15/100
前 0.1%
全站排名 #89
发表论文44 篇
平均评分
年均产出14.7 篇/年
Difan Zou
研究方向
optimization theory · learning theory
19
Turning Internal Gap into Self-Improvement: Promoting the Generation-Understanding Unification in MLLMs
ICLR 2026Poster
通讯29
Learning under Quantization for High-Dimensional Linear Regression
ICLR 2026Poster
三作16
Routing Matters in MoE: Scaling Diffusion Transformers with Explicit Routing Guidance
ICLR 2026Poster
24
A Convergence Analysis of Adaptive Optimizers under Floating-point Quantization
ICLR 2026Poster
三作19
Reshaping Reasoning in LLMs: A Theoretical Analysis of RL Training Dynamics through Pattern Selection
ICLR 2026Poster
三作23
Hyper-SET: Designing Transformers via Hyperspherical Energy Minimization
ICLR 2026Poster
二作23
Does Higher Interpretability Imply Better Utility? A Pairwise Analysis on Sparse Autoencoders
ICLR 2026Poster
通讯23
On the Complexity Theory of Masked Discrete Diffusion: From $\mathrm{poly}(1/\epsilon)$ to Nearly $\epsilon$-Free
ICLR 2026Rejected
5
Physics-Informed Distillation of Diffusion Models for PDE-Constrained Generation
ICLR 2026Withdrawn
二作5
On the Collapse Errors Induced by the Deterministic Sampler for Diffusion Models
ICLR 2026Withdrawn
通讯17
A Random Matrix Analysis of In-context Memorization for Nonlinear Attention
ICLR 2026Rejected
5
Theory of Autoregressive Diffusion Model: Inference Complexity and Conditional Dependency Learning
ICLR 2026Withdrawn
三作18
Understanding the Generalization of Stochastic Gradient Adam in Learning Neural Networks
NeurIPS 2025Poster
通讯17
Hierarchical Koopman Diffusion: Fast Generation with Interpretable Diffusion Trajectory
NeurIPS 2025Poster
三作23
Kernel Regression in Structured Non-IID Settings: Theory and Implications for Denoising Score Learning
NeurIPS 2025Poster
通讯22
On the Robustness of Transformers against Context Hijacking for Linear Classification
NeurIPS 2025Poster
通讯22
Speculative Jacobi-Denoising Decoding for Accelerating Autoregressive Text-to-image Generation
NeurIPS 2025Poster
14
HyPoGen: Optimization-Biased Hypernetworks for Generalizable Policy Generation
ICLR 2025Poster
34
How Does Critical Batch Size Scale in Pre-training?
ICLR 2025Poster
26
How Does Label Noise Gradient Descent Improve Generalization in the Low SNR Regime?
NeurIPS 2025Poster
17
Can Diffusion Models Learn Hidden Inter-Feature Rules Behind Images?
ICML 2025Poster
通讯9
Masked Autoencoders Are Effective Tokenizers for Diffusion Models
ICML 2025Spotlight
22
F-Adapter: Frequency-Adaptive Parameter-Efficient Fine-Tuning in Scientific Machine Learning
NeurIPS 2025Poster
通讯25
On the Feature Learning in Diffusion Models
ICLR 2025Poster
通讯21
Beyond Surface Structure: A Causal Assessment of LLMs' Comprehension ability
ICLR 2025Poster
29
Label Noise Gradient Descent Improves Generalization in the Low SNR Regime
ICLR 2025Rejected
8
On the Robustness of Transformers against Context Hijacking for Linear Classification
ICML 2025Rejected
通讯14
Towards Understanding Fine-Tuning Mechanisms of LLMs via Circuit Analysis
ICML 2025Poster
通讯26
Self-Control of LLM Behaviors by Compressing Suffix Gradient into Prefix Controller
ICLR 2025Rejected
5
Towards a Theoretical Understanding of Memorization in Diffusion Models
ICLR 2025Withdrawn
三作