影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
94.31/100
前 0.3%
全站排名 #197
发表论文35 篇
平均评分
年均产出11.7 篇/年
Hanwang Zhang
研究方向
causal inference · scene graph generation · vision-language
13
Benchmarking Open-Set Recognition Beyond Vision-Language Pre-training
ICLR 2026Rejected
通讯25
Streaming Drag-Oriented Interactive Video Manipulation: Drag Anything, Anytime!
ICLR 2026Poster
通讯23
Reducing Class-Wise Performance Disparity via Margin Regularization
ICLR 2026Poster
通讯18
Real-Time Motion-Controllable Autoregressive Video Diffusion
ICLR 2026Poster
通讯16
Look Carefully: Adaptive Visual Reinforcements in Multimodal Large Language Models for Hallucination Mitigation
ICLR 2026Poster
5
Generative Distribution Distillation
ICLR 2026Withdrawn
16
On Path to Multimodal Generalist: General-Level and General-Bench
ICML 2025Oral
通讯19
Enhancing CLIP Robustness via Cross-Modality Alignment
NeurIPS 2025Spotlight
通讯19
Selftok-Zero: Reinforcement Learning for Visual Generation via Discrete and Autoregressive Visual Tokens
NeurIPS 2025Poster
通讯13
$\mathcal{V}ista\mathcal{DPO}$: Video Hierarchical Spatial-Temporal Direct Preference Optimization for Large Video Models
ICML 2025Poster
22
Co-Reinforcement Learning for Unified Multimodal Understanding and Generation
NeurIPS 2025Spotlight
30
Vinci: Deep Thinking in Text-to-Image Generation using Unified Model with Reinforcement Learning
NeurIPS 2025Poster
通讯8
Towards Semantic Equivalence of Tokenization in Multimodal LLM
ICLR 2025Poster
28
VR-Sampling: Accelerating Flow Generative Model Training with Variance Reduction Sampling
ICLR 2025Withdrawn
三作5
Learning to Animate Images from A Few Videos to Portray Delicate Human Actions
ICLR 2025Withdrawn
11
3D Question Answering via only 2D Vision-Language Models
ICML 2025Poster
15
Geo-3DGS: Multi-view Geometry Consistency for 3D Gaussian Splatting and Surface Reconstruction
ICLR 2025Rejected
通讯9
Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing
ICML 2025Poster
三作5
Object Fusion via Diffusion Time-step for Customized Image Editing with Single Example
ICLR 2025Withdrawn
5
Towards Debiased Source-Free Domain Adaptation
ICLR 2025Withdrawn
三作5
A Closer Look at Time Steps is Worthy of Triple Speed-Up for Diffusion Model Training
ICLR 2025Withdrawn