影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
88.48/100
前 0.7%
全站排名 #438
发表论文34 篇
平均评分
年均产出11.3 篇/年
Zhuoran Yang
研究方向
Large Language Models · Reinforcement Learning
16
Dual-Robust Cross-Domain Offline Reinforcement Learning Against Dynamics Shifts
ICLR 2026Poster
35
Muon Outperforms Adam in Tail-End Associative Memory Learning
ICLR 2026Poster
17
Cross-domain Offline Policy Adaptation with Dynamics- and Value-aligned Data Filtering
ICLR 2026Rejected
34
Taming Polysemanticity in LLMs: Theory-Grounded Feature Recovery via Sparse Autoencoders
ICLR 2026Poster
通讯16
Interpreting Multi-Layer Transformers for In-Context Linear Regression with Varying Covariance
ICLR 2026Rejected
通讯19
How Transformers Learn Causal Structures In-Context: Explainable Mechanism Meets Theoretical Guarantee
ICLR 2026Poster
通讯20
Unlocking Out-of-Distribution Generalization in Transformers via Latent Space Reasoning
ICLR 2026Rejected
通讯5
AV-Odyssey Bench: From Fundamental Audio Perception to Audio-Visual Understanding
ICLR 2026Withdrawn
22
Learning to Incentivize on the Fly: Leader-Follower Games with Policy Recommendation
ICLR 2026Rejected
三作10
Learning in Context, Guided by Choice: A Reward-Free Paradigm for Reinforcement Learning with Transformers
ICLR 2026Rejected
27
On the Mechanism and Dynamics of Modular Addition: Fourier Features, Lottery Ticket, and Grokking
ICLR 2026Rejected
通讯6
Mechanistic Interpretability of In-Context Learning Generalization through Structured Task Curriculum
ICLR 2026Rejected
通讯9
Can VLMs Reason Through Multiple Views?
ICLR 2026Desk Rejected
通讯6
Can Neural Networks Achieve Optimal Computational-statistical Tradeoff? An Analysis on Single-Index Model
ICLR 2025Oral
11
Decoding Rewards in Competitive Games: Inverse Game Theory with Entropy Regularization
ICML 2025Poster
11
In-Context Linear Regression Demystified: Training Dynamics and Mechanistic Interpretability of Multi-Head Softmax Attention
ICML 2025Poster
通讯36
Provable Learning for DEC-POMDPs: Factored Models and Memoryless Agents
ICLR 2025Rejected
三作19
In-Context Reinforcement Learning From Suboptimal Historical Data
ICLR 2025Rejected
15
In-Context Reinforcement Learning From Suboptimal Historical Data
ICML 2025Poster
16
Exploration in the Face of Strategic Responses: Provable Learning of Online Stackelberg Games
ICLR 2025Rejected
通讯11
The Sample Complexity of Online Strategic Decision Making with Information Asymmetry and Knowledge Transportability
ICML 2025Poster
通讯6
STRIDE: A Tool-Assisted LLM Agent Framework for Strategic and Interactive Decision-Making
ICLR 2025Rejected
通讯10
BanditSpec: Adaptive Speculative Decoding via Bandit Algorithms
ICML 2025Poster
通讯15
Quantile-Optimal Policy Learning under Unmeasured Confounding
ICLR 2025Withdrawn
9
An Instrumental Value for Data Production and its Application to Data Pricing
ICML 2025Poster
5
Steer a Crowd: Learning to Persuade a Population in a Stackelberg Game
ICLR 2025Withdrawn
通讯