影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
78.58/100
前 1.5%
全站排名 #970
发表论文26 篇
平均评分
年均产出8.7 篇/年
Shiguang Shan
研究方向
AI safety · large models safety evaluation · Transformers · Foundation models · Self-supervised learning · weakly-supervised learning · Face anti-spoofing · present attack detection · Lip reading · visual speech recognition · Pedestrian re-identification · Deep learning · CNNs · domain adaptation · Facial expression recognition · facial action unit detection · affective computing · Face recognition · face detection · face verification
23
Plan-R1: Safe and Feasible Trajectory Planning as Language Modeling
ICLR 2026Poster
三作16
Uniform Discrete Diffusion with Metric Path for Video Generation
ICLR 2026Poster
23
HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes
ICLR 2026Poster
19
Multi-Level CLIP Transfer for Open-Vocabulary Object Detection
ICLR 2026Rejected
通讯16
Seeing Like Humans: Task-Driven Token Reduction for Accelerated ViT in Robotic Navigation
ICLR 2026Rejected
三作15
Jodi: Unification of Visual Generation and Understanding via Joint Modeling
ICLR 2026Rejected
18
Distribution-Aware Synergistic Evolution for Few-shot Discrimination and Generation
ICLR 2026Rejected
5
A Dual-Branch Disentanglement Diffusion for ID-Attribute Conditional Face Generation
ICLR 2026Withdrawn
三作4
Towards Non-destructive Privacy Protection for LVLMs via node-level localized editing
ICLR 2026Withdrawn
21
ProtInvTree: Deliberate Protein Inverse Folding with Reward-guided Tree Search
NeurIPS 2025Spotlight
24
Autoregressive Video Generation without Vector Quantization
ICLR 2025Poster
30
Revisiting Logit Distributions for Reliable Out-of-Distribution Detection
NeurIPS 2025Poster
29
CtrLoRA: An Extensible and Efficient Framework for Controllable Image Generation
ICLR 2025Poster
三作28
Dysca: A Dynamic and Scalable Benchmark for Evaluating Perception Ability of LVLMs
ICLR 2025Poster
13
un$^2$CLIP: Improving CLIP's Visual Detail Capturing Ability via Inverting unCLIP
NeurIPS 2025Poster
25
Evaluating the Quality of Hallucination Benchmarks for Large Vision-Language Models
ICLR 2025Rejected
9
MATS: An Audio Language Model under Text-only Supervision
ICML 2025Poster
27
Enhancing Zero-shot OOD Detection with Pre-trained Multimodal Foundation Models
NeurIPS 2025Rejected
三作5
Not Only Vision: Evolve Visual Speech Recognition via Peripheral Information
ICLR 2025Withdrawn
三作12
CM^2: Cross-Modal Contextual Modeling for Audio-Visual Speech Enhancement
ICLR 2025Rejected
三作6
Semantic or Covariate? A Study on the Intractable Case of Out-of-Distribution Detection
ICLR 2025Withdrawn
三作