影响力指数
99.26/100
前 0.1%
全站排名 #14
发表论文62
平均评分5.7
年均产出20.7 篇/年

Yilun Du

Researcher@Google·美国·OpenReview
研究方向

Generative Models · Embodied AI / Robot Learning

7.5
15

Reasoning with Sampling: Your Base Model is Smarter Than You Think

ICLR 2026Oral
二作
7.0
22

World-In-World: World Models in a Closed-Loop World

ICLR 2026Oral
6.5
18

Any-Order Flexible Length Masked Diffusion

ICLR 2026Poster
6.5
14

Geometry-aware Policy Imitation

ICLR 2026Poster
6.0
20

Inference-time scaling of diffusion models through classical search

ICLR 2026Poster
通讯
6.0
24

Energy-Based Transformers are Scalable Learners and Thinkers

ICLR 2026Oral
5.5
13

Product of Experts for Visual Generation

ICLR 2026Poster
5.5
24

Self-Improving Loops for Visual Robotic Planning

ICLR 2026Poster
5.3
15

OpenGLA: Topology and Task Adaptive Foundation Model for Power System Graph-Language Answering

ICLR 2026Rejected
5.0
19

Intra-Request Branch Orchestration for Efficient LLM Reasoning

ICLR 2026Rejected
三作
5.0
16

Controllable Video Synthesis via Variational Inference

ICLR 2026Rejected
三作
5.0
18

Implicit State Estimation via Video Replanning

ICLR 2026Rejected
5.0
30

SLM-MUX: Orchestrating Small Language Models for Reasoning

ICLR 2026Poster
通讯
4.5
13

Equilibrium Matching: Generative Modeling with Implicit Energy-Based Models

ICLR 2026Rejected
二作
4.5
24

Model-Based Diffusion Sampling for Predictive Control in Offline Decision Making

ICLR 2026Rejected
三作
4.5
6

TOWARD MEMORY-AIDED WORLD MODELS: BENCHMARKING VIA SPATIAL CONSISTENCY

ICLR 2026Rejected
三作
4.5
16

MoSEL: Modular Self-Reflective Learning for Embodied Decision-Making

ICLR 2026Rejected
三作
4.5
22

Long-Text-to-Image Generation via Compositional Prompt Decomposition

ICLR 2026Poster
三作
4.5
32

MoTVLA: A Vision-Language-Action Model with Unified Fast-Slow Reasoning

ICLR 2026Rejected
4.0
26

Flow Equivariant World Modeling for Partially Observed Dynamic Environments

ICLR 2026Rejected
4.0
17

4D Latent World Model for Robot Planning

ICLR 2026Rejected
通讯
3.5
5

DiZCo: Planning Zero-Shot Coordination in World Models

ICLR 2026Withdrawn
三作
3.5
23

Understanding the Design Space and Cross-Modality Transfer for Vision-Language Models

ICLR 2026Rejected
通讯
3.3
5

Selective Underfitting in Diffusion Models

ICLR 2026Rejected
3.0
5

From Score Distributions to Balance: Plug-and-Play Mixture-of-Experts Routing

ICLR 2026Withdrawn
三作
3.0
20

Test-Time Graph Search for Goal-Conditioned Reinforcement Learning

ICLR 2026Rejected
3.0
7

Building Scalable Real-World Robot Data Generation via Compositional Simulation

ICLR 2026Withdrawn
7.8
22

Generalizable Reasoning through Compositional Energy Minimization

NeurIPS 2025Spotlight
二作
7.8
21

Generative Trajectory Stitching through Diffusion Composition

NeurIPS 2025Spotlight
三作
7.8
17

EvoLM: In Search of Lost Language Model Training Dynamics

NeurIPS 2025Oral
7.5
32

Follow the Energy, Find the Path: Riemannian Metrics from Energy-Based Models

NeurIPS 2025Poster
三作
7.3
19

Grounding Video Models to Actions through Goal Conditioned Exploration

ICLR 2025Spotlight
二作
6.8
25

AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark

ICLR 2025Poster
三作
6.7
20

Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains

ICLR 2025Poster
二作
6.7
19

COMBO: Compositional World Models for Embodied Multi-Agent Cooperation

ICLR 2025Poster
6.6
11

History-Guided Video Diffusion

ICML 2025Poster
6.5
22

Looped Transformers for Length Generalization

ICLR 2025Poster
二作
6.5
18

Multi-Agent Verification: Scaling Test-Time Compute with Multiple Verifiers

COLM 2025Poster
三作
6.4
24

Learning 3D Persistent Embodied World Models

NeurIPS 2025Poster
二作
6.4
15

Compositional Scene Understanding through Inverse Generative Modeling

ICML 2025Poster
三作
6.4
29

Win Fast or Lose Slow: Balancing Speed and Accuracy in Latency-Sensitive Decisions of LLMs

NeurIPS 2025Spotlight
6.4
22

MindJourney: Test-Time Scaling with World Models for Spatial Reasoning

NeurIPS 2025Poster
5.8
33

Solving New Tasks by Adapting Internet Video Knowledge

ICLR 2025Poster
三作
5.5
10

AdaWorld: Learning Adaptable World Models with Latent Actions

ICML 2025Poster
三作
4.3
5

Learning 4D Embodied World Models

ICLR 2025Withdrawn
4.3
5

SnapMem: Snapshot-based 3D Scene Memory for Embodied Exploration and Reasoning

ICLR 2025Withdrawn