影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
87.78/100
前 0.7%
全站排名 #469
发表论文32 篇
平均评分
年均产出10.7 篇/年
Yaowei Wang
研究方向
Machine Learning · Multimedia content understanding · Artificial Intelligence
15
Connecting Where You Look With What You Understand: Trajectory-Driven Localized Understanding for Interactive Vision-Language Models
ICLR 2026Rejected
21
VideoAnchor: Reinforcing Subspace-Structured Visual Cues for Coherent Visual-Spatial Reasoning
ICLR 2026Poster
23
CoPRS: Learning Positional Prior from Chain-of-Thought for Reasoning Segmentation
ICLR 2026Poster
通讯21
CAI: Caption-Sensitive Attention Intervention for Mitigating Object Hallucination in Large Vision-Language Models
ICLR 2026Rejected
22
Norm$\times$Direction: Restoring the Missing Query Norm in Vision Linear Attention
ICLR 2026Withdrawn
21
Rectified Decoupled Dataset Distillation: A Closer Look for Fair and Comprehensive Evaluation
ICLR 2026Poster
通讯14
PIC: Revisiting INR for Image Coding with Fast Encoding and Sub-Millisecond Decoding
ICLR 2026Rejected
17
OVRD: Open-Vocabulary Relation DINO with Text-guided Salient Query Selection
ICLR 2026Rejected
5
In Defense of Prompt-based Continual Learning: Task Interference Mitigation via Confidence-Stratified Classifier Calibration
ICLR 2026Withdrawn
12
One Stone Three Birds: Training-free Core-context-aware Attention for Efficient LLM Prefilling, Decoding, and KV Caching
ICLR 2026Withdrawn
22
Spatial Understanding from Videos: Structured Prompts Meet Simulation Data
NeurIPS 2025Spotlight
16
LoRATv2: Enabling Low-Cost Temporal Modeling in One-Stream Trackers
NeurIPS 2025Spotlight
27
FOCUS: Unified Vision-Language Modeling for Interactive Editing Driven by Referential Segmentation
NeurIPS 2025Poster
9
Perceptually Constrained Precipitation Nowcasting Model
ICML 2025Poster
通讯24
Learning Spatial-Semantic Features for Robust Video Object Segmentation
ICLR 2025Poster
24
An Exploration with Entropy Constrained 3D Gaussians for 2D Video Compression
ICLR 2025Poster
9
Core Context Aware Transformers for Long Context Language Modeling
ICML 2025Poster
7
Learning Fine-Grained Representations through Textual Token Disentanglement in Composed Video Retrieval
ICLR 2025Poster
56
EMMA: Empowering Multi-modal Mamba with Structural and Hierarchical Alignment
ICLR 2025Poster
通讯38
DiffPC: Diffusion-based High Perceptual Fidelity Image Compression with Semantic Refinement
ICLR 2025Poster
4
Building Vision Models upon Heat Conduction
ICLR 2025Withdrawn
29
QFree-Det: Query-Free Detector with Transformer and Sequential Matching
ICLR 2025Rejected
通讯6
MambaVC: Exploring Selective State Spaces for Learned Visual Compression
ICLR 2025Withdrawn
通讯29
Core Context Aware Attention for Long Context Language Modeling
ICLR 2025Rejected
5
VideoUntier: Language-guided Video Feature Disentanglement
ICLR 2025Withdrawn
三作