影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
52.38/100
前 7.4%
全站排名 #4,755
发表论文21 篇
平均评分
年均产出10.5 篇/年
Yuxin Peng
研究方向
cross-media intelligence · image and video analysis and retrieval · computer vision
17
Benchmarking Large Vision-Language Models on Fine-Grained Image Tasks: A Comprehensive Evaluation
ICLR 2026Poster
二作11
Fine-R1: Make Multi-modal LLMs Excel in Fine-Grained Visual Recognition by Chain-of-Thought Reasoning
ICLR 2026Poster
三作12
E-CommerceVideo: A Benchmark and approach for E-Commerce Video Generation from product Images
ICLR 2026Rejected
通讯12
Know but can't Say: Exploring the Hidden Knowledge of Large Vision-Language Models for Fine-grained Perception
ICLR 2026Rejected
三作15
SiTu: A Simple Training-Free Thinking-with-Image Approach via Uncertainty Guidance
ICLR 2026Rejected
三作5
PR-VTON: Enhancing Detail Fidelity in Virtual Try-On with Refined Positional Encoding
ICLR 2026Withdrawn
通讯5
MARS: Mamba-driven Adaptive Reordering Scheme for Semantic Occupancy Prediction in Autonomous Driving
ICLR 2026Rejected
二作5
RCR: Relation-Centric Reasoning with Large Language Models for Knowledge-based Question Answering
ICLR 2026Rejected
二作6
Part-Aware CLIP: Enhancing Fine-Grained Understanding with Part-level Descriptions
ICLR 2026Rejected
二作6
FOE-RL: Flexible Online Reinforcement Learning for Efficient Inference in Large Language Models
ICLR 2026Rejected
二作15
MAI: A Multi-turn Aggregation-Iteration Model for Composed Image Retrieval
ICLR 2025Poster
通讯11
Balancing Preservation and Modification: A Region and Semantic Aware Metric for Instruction-Based Image Editing
ICML 2025Poster
三作14
Analyzing and Boosting the Power of Fine-Grained Visual Recognition for Multi-modal Large Language Models
ICLR 2025Poster
通讯15
Investigating Domain Gaps for Indoor 3D Object Detection
ICLR 2025Rejected
18
VISCON: Identifying and Benchmarking Vision Hallucination for Large Vision-Language Model
ICLR 2025Rejected
二作7
BiC-Occ: Bi-directional Circulated 3D Occupancy Prediction for Autonomous Driving
ICLR 2025Rejected
二作4
Pose-guided Motion Diffusion Model for Text-to-motion Generation
ICLR 2025Withdrawn
7
Pseudo Meets Zero: Boosting Zero-Shot Composed Image Retrieval with Synthetic Images
ICLR 2025Rejected
通讯5
Contextually Harmonious Local Video Editing
ICLR 2025Withdrawn
三作7
Ger: Generation, Evaluation and Reflection Enhanced LLM for Knowledge Graph Question Answering
ICLR 2025Rejected
通讯5
Balancing Differential Discriminative Knowledge For Clothing-Irrelevant Lifelong Person Re-identification
ICLR 2025Withdrawn
三作