影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
93.59/100
前 0.3%
全站排名 #221
发表论文34 篇
平均评分
年均产出11.3 篇/年
Ranjay Krishna
研究方向
vision and language · scene graphs · human understanding
21
Unfolding Spatial Cognition: Evaluating Multimodal Models on Visual Simulations
ICLR 2026Poster
通讯25
AdaReasoner: Dynamic Tool Orchestration for Iterative Visual Reasoning
ICLR 2026Poster
22
Theory of Space: Can Foundation Models Construct Spatial Beliefs through Active Exploration?
ICLR 2026Poster
29
MMMG: A Comprehensive and Reliable Benchmark for Multitask Multimodal Generation
ICLR 2026Rejected
24
Generate Any Scene: Scene Graph Driven Data Synthesis for Visual Generation Training
ICLR 2026Poster
通讯36
ThinkMorph: Emergent Properties in Multimodal Interleaved Chain-of-Thought Reasoning
ICLR 2026Poster
28
Spatial Mental Modeling from Limited Views
ICLR 2026Poster
25
TrustGen: A Platform of Dynamic Benchmarking on the Trustworthiness of Generative Foundation Models
ICLR 2026Poster
5
PointArena: Probing Multimodal Grounding Through Language-Guided Pointing
ICLR 2026Withdrawn
通讯14
Spurious Rewards: Rethinking Training Signals in RLVR
ICLR 2026Desk Rejected
25
MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine
ICLR 2026Rejected
通讯24
The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains
COLM 2025Poster
22
Convergent Functions, Divergent Forms
NeurIPS 2025Poster
通讯21
VAGEN: Reinforcing World Model Reasoning for Multi-Turn VLM Agents
NeurIPS 2025Poster
20
Interleaved Scene Graphs for Interleaved Text-and-Image Generation Assessment
ICLR 2025Spotlight
通讯9
SAM2Act: Integrating Visual Foundation Model with A Memory Architecture for Robotic Manipulation
ICML 2025Poster
11
SAT: Dynamic Spatial Aptitude Training for Multimodal Language Models
COLM 2025Poster
37
AHA: A Vision-Language-Model for Detecting and Reasoning Over Failures in Robotic Manipulation
ICLR 2025Poster
28
Visual Representations inside the Language Model
COLM 2025Poster
通讯57
Language Model Preference Evaluation with Multiple Weak Evaluators
ICLR 2025Rejected
通讯23
Diffusion Models are Few-shot Learners for Dense Vision Tasks
ICLR 2025Rejected
通讯23
FocalLens: Instruction Tuning Enables Zero-Shot Conditional Image Representations
ICLR 2025Rejected
4
Coarse Correspondences Boost 3D Spacetime Understanding in Multimodal Language Model
ICLR 2025Withdrawn
通讯