影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
89.29/100
前 0.6%
全站排名 #407
发表论文26 篇
平均评分
年均产出8.7 篇/年
Antonio Orvieto
Principal Researcher@ELLIS Institute Tübingen, Max Planck Institute for Intelligent Systems, Tübingen AI Center, Tübingen, Germany·德国·OpenReview
研究方向
Sequence Models · Genomic Language Models · State-Space Models · Theory of Deep Learning · Convex Optimization · Nonconvex Optimization · Deep Learning · Stochastic Optimization
15
How does the optimizer implicitly bias the model merging loss landscape?
ICLR 2026Poster
11
Selective Rotary Position Embedding
ICLR 2026Poster
16
On the Interaction of Batch Noise, Adaptivity, and Compression, under $(L_0,L_1)$-Smoothness: An SDE Approach
ICLR 2026Rejected
16
Design Principles for Sequence Models via Coefficient Dynamics
ICLR 2026Rejected
二作16
Improved state mixing in higher-order and block diagonal linear recurrent networks
ICLR 2026Rejected
二作20
Revisiting Associative Recall in Modern Recurrent Models
ICLR 2026Rejected
二作15
Scaling Behavior of Discrete Diffusion Language Models
ICLR 2026Poster
通讯11
Is your batch size the problem? Revisiting the Adam-SGD gap in language modeling
ICLR 2026Rejected
三作14
Fixed-Point RNNs: Interpolating from Diagonal to Dense
NeurIPS 2025Spotlight
通讯18
In Search of Adam’s Secret Sauce
NeurIPS 2025Oral
一作25
Geometric Inductive Biases of Deep Networks: The Role of Data and Architecture
ICLR 2025Spotlight
二作23
Generalized Linear Mode Connectivity for Transformers
NeurIPS 2025Oral
22
Adaptive Methods through the Lens of SDEs: Theoretical Insights on the Role of Noise
ICLR 2025Poster
24
When recalling in-context, Transformers are not SSMs
NeurIPS 2025Rejected
二作14
When, Where and Why to Average Weights?
ICML 2025Poster
二作21
Enhancing Optimizer Stability: Momentum Adaptation of NGN Step-size
ICLR 2025Rejected
三作23
Enhancing Optimizer Stability: Momentum Adaptation of The NGN Step-size
NeurIPS 2025Poster
三作11
Generalized Interpolating Discrete Diffusion
ICML 2025Poster
11
NIMBA : Towards Robust and Principled Processing of Point Clouds With SSMs
ICLR 2025Rejected
三作