影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
63.69/100
前 3.8%
全站排名 #2,461
发表论文11 篇
平均评分
年均产出3.7 篇/年
Afshin Dehghan
研究方向
Multimodal Understanding and Reasoning · Multimodal foundation model · 3D Scene Understanding · Perception
23
Rooms from Motion: Un-posed Indoor 3D Object Detection as Localization and Mapping
NeurIPS 2025Poster
三作23
MM1.5: Methods, Analysis & Insights from Multimodal LLM Fine-tuning
ICLR 2025Poster
10
Understanding Alignment in Multimodal LLMs: A Comprehensive Study
ICLR 2025Rejected
13
FlexTok: Resampling Images into 1D Token Sequences of Flexible Length
ICML 2025Poster
通讯24
StreamBridge: Turning Your Offline Video Large Language Model into a Proactive Streaming Assistant
NeurIPS 2025Poster
17
SlowFast-LLaVA-1.5: A Family of Token-Efficient Video Large Language Models for Long-Form Video Understanding
COLM 2025Poster
通讯27
UniGen: Enhanced Training & Test-Time Strategies for Unified Multimodal Understanding and Generation
NeurIPS 2025Poster
通讯21
SlowFast-LLaVA: A strong training-free baseline for video large language models
ICLR 2025Rejected
通讯