影响力指数
92.07/100
前 0.4%
全站排名 #278
发表论文46
平均评分5.2
年均产出15.3 篇/年

Liang Wang

Full Professor@Institute of Automation, CAS,China·中国·OpenReview
研究方向

action recognition · action detection · multimodal analysis · vision and language · detection · segmentation · tracking

7.0
30

AVoCaDO: An Audiovisual Video Captioner Driven by Temporal Orchestration

ICLR 2026Poster
通讯
6.5
16

R1-Reward: Training Multimodal Reward Model Through Stable Reinforcement Learning

ICLR 2026Poster
通讯
6.0
14

Variational Reasoning for Language Models

ICLR 2026Poster
6.0
16

Thyme: Think Beyond Images

ICLR 2026Poster
5.5
15

PlotCraft: Pushing the Limits of LLMs for Complex and Interactive Data Visualization

ICLR 2026Withdrawn
5.5
20

BaseReward: A Strong Baseline for Multimodal Reward Model

ICLR 2026Poster
通讯
5.0
43

MME-Unify: A Comprehensive Benchmark for Unified Multimodal Understanding and Generation Models

ICLR 2026Poster
通讯
5.0
42

ToolWeaver: Weaving Collaborative Semantics for Scalable Tool Use in Large Language Models

ICLR 2026Poster
通讯
4.7
30

VidBridge-R1: Bridging QA and Captioning for RL-based Video Understanding Models with Intermediate Proxy Tasks

ICLR 2026Poster
通讯
4.5
13

OpenGPT-4o-Image: A Comprehensive Dataset for Advanced Image Generation and Editing

ICLR 2026Rejected
4.5
21

UAOR: Uncertainty-aware Observation Reinjection for Vision-Language-Action Models

ICLR 2026Withdrawn
通讯
4.4
14

C$^3$-Bench: Evaluating and Achieving Controllable Code Completion in Code LLM

ICLR 2026Withdrawn
4.0
20

Real-Time Evaluation for Novel Class Discovery at Test Time

ICLR 2026Rejected
通讯
4.0
18

Reinforcing General Reasoning Without Verifiers

ICLR 2026Poster
4.0
5

Conformal Uncertainty Indicator for Continual Test-Time Adaptation

ICLR 2026Withdrawn
通讯
4.0
45

From Genomic Whispers to Therapeutics: Multi-Resolution Transcriptome-Guided Diffusion Models for Drug Design and Screening

ICLR 2026Withdrawn
通讯
4.0
22

Binding Mode Matters: Residue-Guided Drug Discovery via Explorative Preferences

ICLR 2026Rejected
通讯
4.0
20

Uncovering Competing Poisoning Attacks in Retrieval-Augmented Generation

ICLR 2026Rejected
通讯
4.0
19

BridgeV2W: Bridging Video Generation Models to Embodied World Models via Embodiment Masks

ICLR 2026Withdrawn
通讯
3.5
14

Latent Sketchpad: Sketching Visual Thoughts to Elicit Multimodal Reasoning in MLLMs

ICLR 2026Rejected
3.5
5

EgoDemoGen: Novel Egocentric Demonstration Generation Enables Viewpoint-Robust Manipulation

ICLR 2026Withdrawn
通讯
6.8
24

DAA: Amplifying Unknown Discrepancy for Test-Time Discovery

NeurIPS 2025Poster
通讯
6.8
23

Reinforcing Spatial Reasoning in Vision-Language Models with Interwoven Thinking and Visual Drawing

NeurIPS 2025Poster
6.8
23

GOOD: Training-Free Guided Diffusion Sampling for Out-of-Distribution Detection

NeurIPS 2025Poster
6.8
34

MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans?

ICLR 2025Poster
6.3
23

MolSpectra: Pre-training 3D Molecular Representation with Multi-modal Energy Spectra

ICLR 2025Poster
通讯
6.2
25

Integrating Protein Dynamics into Structure-Based Drug Design via Full-Atom Stochastic Flows

ICLR 2025Poster
6.1
11

MM-RLHF: The Next Step Forward in Multimodal LLM Alignment

ICML 2025Poster
6.0
24

Graffe: Graph Representation Learning Enabled via Diffusion Probabilistic Models

ICLR 2025Rejected
5.5
24

BridgeVLA: Input-Output Alignment for Efficient 3D Manipulation Learning with Vision-Language Models

NeurIPS 2025Poster
5.0
6

Hierarchical Multimodal Knowledge Matching for Training-Free Open-Vocabulary Object Detection

ICLR 2025Withdrawn
三作
5.0
28

TimeRAF: Retrieval-Augmented Foundation model for Zero-shot Time Series Forecasting

ICLR 2025Withdrawn
5.0
14

CONSTRAINT-AWARE ZERO-SHOT VISION-LANGUAGE NAVIGATION IN CONTINUOUS ENVIRONMENTS

ICLR 2025Withdrawn
通讯
4.8
6

Dual Flows with Contrastive Guidance for Generating Highly Designable Proteins

ICLR 2025Rejected
4.7
8

Beyond Filtering: Adaptive Image-Text Quality Enhancement for MLLM Pretraining

ICLR 2025Withdrawn
通讯
4.3
17

TUI: A Conformal Uncertainty Indicator for Continual Test-Time Adaptation

ICLR 2025Withdrawn
通讯
4.3
27

Controllable Continual Test-Time Adaptation

ICLR 2025Withdrawn
通讯