影响力指数
85.4/100
前 0.9%
全站排名 #578
发表论文28
平均评分5.1
年均产出9.3 篇/年

Zhuotao Tian

Full Professor@Harbin Institute of Technology (Shenzhen)·中国·OpenReview
研究方向

Multi-modal perception · Large Language Models · Multi-modal understanding and analysis · Few-shot segmentation and Few-shot Learning · Semantic Segmentation · Computer Vision

7.0
42

Efficient Reasoning with Balanced Thinking

ICLR 2026Poster
通讯
6.5
16

PointRePar : SpatioTemporal Point Relation Parsing for Robust Category-Unified 3D Tracking

ICLR 2026Poster
三作
6.0
24

SafeMVDrive: Multi-view Safety-Critical Driving Video Generation in the Real World Domain

ICLR 2026Withdrawn
三作
6.0
22

Dynamic-dLLM: Dynamic Cache-Budget and Adaptive Parallel Decoding for Training-Free Acceleration of Diffusion LLM

ICLR 2026Poster
通讯
5.5
36

FlashVID: Efficient Video Large Language Models via Training-free Tree-based Spatiotemporal Token Merging

ICLR 2026Oral
通讯
4.8
40

Plug-and-Play Fidelity Optimization for Diffusion Transformer Acceleration via Cumulative Error Minimization

ICLR 2026Poster
4.5
5

Video-ToC: Video Tree-of-Cue Reasoning

ICLR 2026Withdrawn
二作
4.5
32

Uni-DPO: A Unified Paradigm for Dynamic Preference Optimization of LLMs

ICLR 2026Poster
三作
4.5
18

GRASP-GS: Geometric Registration and Dual-Stag Saliency Pruning for Efficient 3D Gaussian Splatting

ICLR 2026Rejected
通讯
4.5
6

ES-GGT: Efficient Submap-based Visual Geometry Grounded Transformer with Spatial Memory Alignment

ICLR 2026Rejected
4.5
26

Multimodal Dataset Distillation via Phased Teacher Models

ICLR 2026Poster
通讯
4.5
26

LongHorizonUI: A Unified Framework for Robust long-horizon Task Automation of GUI Agent

ICLR 2026Poster
通讯
4.0
5

Memory Forcing: Spatio-Temporal Memory for Consistent Scene Generation on Minecraft

ICLR 2026Withdrawn
4.0
4

Consistency Beyond Contrast: Enhancing Object Detection Robustness through Scene-Augmented Feature Alignment

ICLR 2026Withdrawn
4.0
5

VisionFocus: Towards Efficient Hallucination Mitigation via Token-Aware Visual Enhancement

ICLR 2026Withdrawn
通讯
4.0
6

PrecogUI: Proactive GUI Agents via Pre-cognitive Simulation and Experience Retrieval

ICLR 2026Withdrawn
通讯
3.5
5

Amodal SAM: Open-World Amodal Segmentation

ICLR 2026Withdrawn
二作