影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
73.27/100
前 2.1%
全站排名 #1,375
发表论文25 篇
平均评分
年均产出8.3 篇/年
Xilin Chen
研究方向
Scene Analysis · Scene Understanding · Vision and Language · Face Perception · Sign Language
23
Plan-R1: Safe and Feasible Trajectory Planning as Language Modeling
ICLR 2026Poster
通讯23
HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes
ICLR 2026Poster
通讯16
Seeing Like Humans: Task-Driven Token Reduction for Accelerated ViT in Robotic Navigation
ICLR 2026Rejected
通讯25
LexSign: Learning Sign Language from Lexical Descriptions
ICLR 2026Rejected
通讯15
Jodi: Unification of Visual Generation and Understanding via Joint Modeling
ICLR 2026Rejected
通讯5
MsTok: Query-Based Multi-Scale 1D Visual Tokenization
ICLR 2026Withdrawn
通讯18
Distribution-Aware Synergistic Evolution for Few-shot Discrimination and Generation
ICLR 2026Rejected
通讯5
CFMAE: A Coarse-to-Fine Vision Pre-training Framework for Hierarchical Representation Learning
ICLR 2026Withdrawn
通讯16
MoTE: Mixture of Ternary Experts for Memory-efficient Large Multimodal Models
ICLR 2026Rejected
通讯4
Towards Non-destructive Privacy Protection for LVLMs via node-level localized editing
ICLR 2026Withdrawn
通讯21
ProtInvTree: Deliberate Protein Inverse Folding with Reward-guided Tree Search
NeurIPS 2025Spotlight
通讯30
Revisiting Logit Distributions for Reliable Out-of-Distribution Detection
NeurIPS 2025Poster
通讯29
CtrLoRA: An Extensible and Efficient Framework for Controllable Image Generation
ICLR 2025Poster
通讯28
Dysca: A Dynamic and Scalable Benchmark for Evaluating Perception Ability of LVLMs
ICLR 2025Poster
通讯13
un$^2$CLIP: Improving CLIP's Visual Detail Capturing Ability via Inverting unCLIP
NeurIPS 2025Poster
通讯25
Evaluating the Quality of Hallucination Benchmarks for Large Vision-Language Models
ICLR 2025Rejected
通讯9
MATS: An Audio Language Model under Text-only Supervision
ICML 2025Poster
通讯27
Enhancing Zero-shot OOD Detection with Pre-trained Multimodal Foundation Models
NeurIPS 2025Rejected
通讯5
Not Only Vision: Evolve Visual Speech Recognition via Peripheral Information
ICLR 2025Withdrawn
通讯22
Robotic Programmer: Video Instructed Policy Code Generation for Robotic Manipulation
ICLR 2025Rejected
通讯5
R2C: Mapping Room to Chessboard to Unlock LLM As Low-Level Action Planner
ICLR 2025Withdrawn
通讯5
M4U: Evaluating Multilingual Understanding and Reasoning for Large Multimodal Models
ICLR 2025Withdrawn
通讯12
CM^2: Cross-Modal Contextual Modeling for Audio-Visual Speech Enhancement
ICLR 2025Rejected
通讯6
Semantic or Covariate? A Study on the Intractable Case of Out-of-Distribution Detection
ICLR 2025Withdrawn
通讯