影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
89.49/100
前 0.6%
全站排名 #398
发表论文44 篇
平均评分
年均产出14.7 篇/年
Ying Shan
研究方向
Multi-modal understanding and generation
15
From Prediction to Perfection: Introducing Refinement to Autoregressive Image Generation
ICLR 2026Poster
通讯14
IC-Custom: Diverse Image Customization via In-Context Learning
ICLR 2026Poster
通讯17
AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation
ICLR 2026Rejected
29
Rolling Forcing: Autoregressive Long Video Diffusion in Real Time
ICLR 2026Poster
19
GenCompositor: Generative Video Compositing with Diffusion Transformer
ICLR 2026Poster
16
ToonComposer: Streamlining Cartoon Production with Generative Post-Keyframing
ICLR 2026Poster
通讯4
MeshWeaver: Sparse-Voxel-Guided Surface Weaving for Autoregressive Mesh Generation
ICLR 2026Withdrawn
三作12
From Denoising to Refining: A Corrective Framework for Vision-Language Diffusion Model
ICLR 2026Rejected
22
DVT-LLaVA: Vision-Language Model Personalization with Disentangled Visual Tuning
ICLR 2026Rejected
通讯13
Aligning Latent Spaces with Flow Priors
ICLR 2026Rejected
9
GRPO-CARE: Consistency-Aware Reinforcement Learning for Multimodal Reasoning
ICLR 2026Withdrawn
5
AudioStory: Generating Long-Form Narrative Audio with Large Language Models
ICLR 2026Withdrawn
通讯5
Unified Single Transformer for Multimodal Video Understanding and Generation
ICLR 2026Withdrawn
通讯11
Fourier Minds, Forget Less: Discrete Fourier Transform for Fast and Robust Continual Learning in LLMs
ICLR 2026Withdrawn
11
LoRA-Gen: Specializing Large Language Model via Online LoRA Generation
ICML 2025Poster
通讯25
UniPixel: Unified Object Referring and Segmentation for Pixel-Level Visual Reasoning
NeurIPS 2025Poster
26
MindOmni: Unleashing Reasoning Generation in Vision Language Models with RGPO
NeurIPS 2025Poster
通讯10
HaploVL: A Single-Transformer Baseline for Multi-Modal Understanding
ICML 2025Poster
14
Taming Rectified Flow for Inversion and Editing
ICML 2025Poster
通讯22
LoRA-Gen: Specializing Language Model via Online LoRA Generation
ICLR 2025Withdrawn
通讯28
FreeSplatter: Pose-free Gaussian Splatting for Sparse-view 3D Reconstruction
ICLR 2025Rejected
三作6
Mani-GS: Gaussian Splatting Manipulation with Triangular Mesh
ICLR 2025Withdrawn
31
PPLLaVA: Varied Video Sequence Understanding With Prompt Guidance
ICLR 2025Rejected
7
SEED-X: Multimodal Models in Real World
ICLR 2025Rejected
通讯6
SEED-Story: Multimodal Long Story Generation with Large Language Model
ICLR 2025Withdrawn
6
Self-Conditioned Diffusion Model for Consistent Human Image and Video Synthesis
ICLR 2025Withdrawn
6
GPT4LoRA: Optimizing LoRA Combination via MLLM Self-Reflection
ICLR 2025Rejected
通讯