影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
78.33/100
前 1.5%
全站排名 #989
发表论文25 篇
平均评分
年均产出12.5 篇/年
Fahad Shahbaz Khan
研究方向
Multimodal learning · Object tracking · Action Recognition · Object detection · object recognition · scene understanding · image analysis
11
TerraFM: A Scalable Foundation Model for Unified Multisensor Earth Observation
ICLR 2026Poster
20
Agent-X: Evaluating Deep Multimodal Reasoning in Vision-Centric Agentic Tasks
ICLR 2026Poster
28
VideoMathQA: Benchmarking Mathematical Reasoning via Multimodal Understanding in Video
ICLR 2026Poster
通讯15
StageVAR: Stage-Aware Acceleration for Visual Autoregressive Models
ICLR 2026Rejected
18
MediX-R1: Open Ended Medical Reinforcement Learning
ICLR 2026Rejected
5
RainDiff: End to End Precipitation Nowcasting Via Token-wise Attention Diffusion
ICLR 2026Withdrawn
三作6
MAVOS-DD: Multilingual Audio-Video Open-Set Deepfake Detection Benchmark
ICLR 2026Withdrawn
4
ThinkGeo: Evaluating Tool-Augmented Agents for Remote Sensing Tasks
ICLR 2026Withdrawn
5
Towards Multimodal Understanding, Reasoning, and Tool Usage across Vision, Speech, and Audio in Long Videos
ICLR 2026Withdrawn
7
GeoVLM-R1: Reinforcement Fine-Tuning for Improved Remote Sensing Reasoning
ICLR 2026Withdrawn
8
VideoMolmo: Spatio-Temporal Grounding meets Pointing
ICLR 2026Withdrawn
-1
MATRIX: Multimodal Agent Tuning for Robust Tool-Use Reasoning
ICLR 2026Withdrawn
16
Open-YOLO 3D: Towards Fast and Accurate Open-Vocabulary 3D Instance Segmentation
ICLR 2025Oral
通讯24
DEFT: Decompositional Efficient Fine-Tuning for Text-to-Image Models
NeurIPS 2025Poster
三作29
One-Prompt-One-Story: Free-Lunch Consistent Text-to-Image Generation Using a Single Prompt
ICLR 2025Spotlight
24
ZeroDiff: Solidified Visual-semantic Correlation in Zero-Shot Learning
ICLR 2025Poster
22
AdaIR: Adaptive All-in-One Image Restoration via Frequency Mining and Modulation
ICLR 2025Poster
通讯13
GeoPixel: Pixel Grounding Large Multimodal Model in Remote Sensing
ICML 2025Poster
40
$InterLCM$: Low-Quality Images as Intermediate States of Latent Consistency Models for Effective Blind Face Restoration
ICLR 2025Poster
10
GenZSL: Generative Zero-Shot Learning Via Inductive Variational Autoencoder
ICML 2025Poster
通讯5
Dynamic Pre-training: Towards Efficient and Scalable All-in-One Image Restoration
ICLR 2025Withdrawn
5
How Good is my Video LMM? Complex Video Reasoning and Robustness Evaluation Suite for Video-LMMs
ICLR 2025Withdrawn
5
Induction Rather Than Imagination: Generative Zero-Shot Learning Via Inductive Variational Autoencoder
ICLR 2025Withdrawn
通讯6
VideoGPT+: Integrating Image and Video Encoders for Enhanced Video Understanding
ICLR 2025Withdrawn
通讯4
GroupMamba: Parameter-Efficient and Accurate Group Visual State Space Model
ICLR 2025Withdrawn
通讯