影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
89.41/100
前 0.6%
全站排名 #403
发表论文40 篇
平均评分
年均产出13.3 篇/年
Yutaka Matsuo
研究方向
deep learning · web mining · social media
16
Does “Do Differentiable Simulators Give Better Policy Gradients?” Give Better Policy Gradients?
ICLR 2026Poster
通讯17
Quantization-Aware Diffusion Models For Maximum Likelihood Training
ICLR 2026Poster
三作13
C-Voting: Confidence-Based Test-Time Voting without Explicit Energy Functions
ICLR 2026Poster
通讯23
SELF-HARMONY: LEARNING TO HARMONIZE SELF-SUPERVISION AND SELF-PLAY IN TEST-TIME REINFORCEMENT LEARNING
ICLR 2026Poster
19
RL Squeezes, SFT Expands: A Comparative Study of Reasoning LLMs
ICLR 2026Poster
通讯23
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying
ICLR 2026Rejected
通讯20
Geometry of Nash Mirror Dynamics: Adaptive $\beta$-Control for Stable and Bias-Robust Self-Improving LLM Agents
ICLR 2026Rejected
通讯14
Unlocking Noise-Resistant Vision: Key Architectural Secrets for Robust Models
ICLR 2026Rejected
通讯13
WorldPack: Compressed Memory Improves Spatial Consistency in Video World Modeling
ICLR 2026Rejected
30
Leave No Observation Behind: Real-time Correction for VLA Action Chunks
ICLR 2026Rejected
29
Recurrent model for Sequential reasoning
ICLR 2026Rejected
通讯5
Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback
ICLR 2026Withdrawn
25
Emergent Misalignment from Superposition
ICLR 2026Withdrawn
通讯10
Vertical Attention: Automatic Exploration of Inter-Layer Connections in Transformer-based Language Models
ICLR 2026Rejected
通讯19
Provably Efficient RL under Episode-Wise Safety in Constrained MDPs with Linear Function Approximation
NeurIPS 2025Spotlight
通讯24
Inference-Time Text-to-Video Alignment with Diffusion Latent Beam Search
NeurIPS 2025Poster
三作20
Topology of Reasoning: Understanding Large Reasoning Models through Reasoning Graph Properties
NeurIPS 2025Poster
通讯15
Beyond Induction Heads: In-Context Meta Learning Induces Multi-Phase Circuit Emergence
ICML 2025Poster
通讯24
Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form
ICLR 2025Poster
通讯40
Rethinking Evaluation of Sparse Autoencoders through the Representation of Polysemous Words
ICLR 2025Poster
通讯27
CityNav: Language-Goal Aerial Navigation Dataset Using Geographic Information
ICLR 2025Rejected
26
Bridging Lottery Ticket and Grokking: Understanding Grokking from Inner Structure of Networks
ICLR 2025Rejected
三作6
FullDiffusion: Diffusion Models Without Time Truncation
ICLR 2025Rejected
通讯10
ToM-agent: Large Language Models as Theory of Mind Aware Generative Agents with Counterfactual Reflection
ICLR 2025Rejected
通讯21
The Geometry of Phase Transitions in Diffusion Models: Tubular Neighbourhoods and Singularities
ICLR 2025Rejected
通讯6
MMA: Benchmarking Multi-Modal Large Language Model in Ambiguity Contexts
ICLR 2025Withdrawn
14
RAGDP: Retrieve-Augmented Generative Diffusion Policy
ICLR 2025Rejected
通讯9
Maximum Likelihood Estimation for Flow Matching by Direct Second-order Trace Objective
ICLR 2025Rejected
三作27
Curse of Instructions: Large Language Models Cannot Follow Multiple Instructions at Once
ICLR 2025Rejected
通讯