影响力指数
89/100
前 0.7%
全站排名 #420
发表论文56
平均评分5.3
年均产出18.7 篇/年

Pengfei Wan

Director@Kuaishou Technology·中国·OpenReview
研究方向

deep learning · computer vision

7.0
30

AVoCaDO: An Audiovisual Video Captioner Driven by Temporal Orchestration

ICLR 2026Poster
6.5
14

Latent Diffusion Model without Variational Autoencoder

ICLR 2026Poster
6.0
21

UniVideo: Unified Understanding, Generation, and Editing for Videos

ICLR 2026Poster
6.0
15

Unified In-Context Video Editing

ICLR 2026Poster
6.0
20

Easier Painting Than Thinking: Can Text-to-Image Models Set the Stage, but Not Direct the Play?

ICLR 2026Poster
5.5
16

Mitigating Noise Shift in Denoising Generative Models with Noise Awareness Guidance

ICLR 2026Poster
5.5
12

VMoBA: Mixture-of-Block Attention for Video Diffusion Models

ICLR 2026Poster
5.5
18

Improving Autoregressive Video Modeling with History Understanding

ICLR 2026Poster
5.5
24

SimpleGVR: A Simple Baseline for Latent-Cascaded Generative Video Super-Resolution

ICLR 2026Poster
5.5
33

VideoSearch Reasoner: Boosting Multimodal Reward Models through Think with Image Reasoning

ICLR 2026Rejected
5.5
20

From Inpainting to Editing: A Self-Bootstrapping Paradigm for Context-Rich Visual Dubbing

ICLR 2026Rejected
5.5
22

AdaViewPlanner: Adapting Video Diffusion Models for Viewpoint Planning in 4D Scenes

ICLR 2026Poster
5.3
13

Free Lunch Alignment of Text-to-Image Diffusion Models without Preference Image Pairs

ICLR 2026Rejected
5.3
19

Learning Video Generation for Robotic Manipulation with Collaborative Trajectory Control

ICLR 2026Poster
5.3
15

DiffMoE: Dynamic Token Selection for Scalable Diffusion Transformers

ICLR 2026Rejected
5.0
14

Astra: General Interactive World Model with Autoregressive Denoising

ICLR 2026Poster
5.0
11

RelightMaster: Precise Video Relighting with Multi-plane Light Images

ICLR 2026Rejected
4.7
19

A Guide to Training Consistency Models

ICLR 2026Rejected
三作
4.7
4

A Reason-then-Describe Instruction Interpreter for Controllable Video Generation

ICLR 2026Withdrawn
4.7
30

VidBridge-R1: Bridging QA and Captioning for RL-based Video Understanding Models with Intermediate Proxy Tasks

ICLR 2026Poster
4.5
26

Scaling Image and Video Generation via Test-Time Evolutionary Search

ICLR 2026Rejected
4.5
20

TivTok: Broadcasting Time-Invariant Tokens for Scalable Video Tokenization

ICLR 2026Rejected
4.5
20

FilMaster: Bridging Cinematic Principles and Generative AI for Automated Film Generation

ICLR 2026Poster
4.5
5

OmniX: From Unified Panoramic Generation and Perception to Graphics-Ready 3D Scenes

ICLR 2026Withdrawn
4.5
5

Interpreting Any Condition to Caption for Controllable Video Generation

ICLR 2026Withdrawn
4.5
27

UniMMVSR: A Unified Multi-Modal Framework for Cascaded Video Super-Resolution

ICLR 2026Rejected
4.5
5

CamPilot: Improving Camera Control in Video Diffusion Model with Efficient Camera Reward Feedback

ICLR 2026Withdrawn
4.5
5

Efficient Training-Free High-Resolution Synthesis with Energy Rectification in Diffusion Models

ICLR 2026Withdrawn
4.5
13

OpenGPT-4o-Image: A Comprehensive Dataset for Advanced Image Generation and Editing

ICLR 2026Rejected
4.0
5

PhysMaster: Mastering Physical Representation for Video Generation via Reinforcement Learning

ICLR 2026Withdrawn
4.0
16

SPF-Portrait: Towards Pure Text-to-Portrait Customization with Semantic Pollution-Free Fine-Tuning

ICLR 2026Rejected
4.0
4

Score Augmentation for Diffusion Models

ICLR 2026Withdrawn
4.0
5

VideoCanvas: Unified Video Completion from Arbitrary Spatiotemporal Patches via In-Context Conditioning

ICLR 2026Withdrawn
4.0
5

RealUnify: Do Unified Models Truly Benefit from Unification? A Comprehensive Benchmark

ICLR 2026Withdrawn
3.5
10

Terra: Explorable Native 3D World Model with Point Latents

ICLR 2026Withdrawn
3.5
5

VFXMaster: Unlocking Dynamic Visual Effect Generation via In-Context Learning

ICLR 2026Withdrawn
3.3
4

EmoDialogCN: A Multimodal Mandarin Dyadic Dialogue Dataset of Emotions

ICLR 2026Withdrawn
通讯
3.0
5

Bowtie-flow: Efficient High-Resolution Video Generation with Prior Preservation

ICLR 2026Withdrawn
2.5
12

Fast to Train, Fast to Sample: Stable Velocity for Flow Matching

ICLR 2026Rejected