影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
68.97/100
前 2.8%
全站排名 #1,826
发表论文25 篇
平均评分
年均产出8.3 篇/年
Sarah Monazam Erfani
研究方向
AI safety · adversarial machine learning · unsupervised learning · anomaly detection
13
On the Bayes Inconsistency of Disagreement Discrepancy Surrogates
ICLR 2026Poster
通讯20
Toward Universal and Transferable Jailbreak Attacks on Vision-Language Models
ICLR 2026Poster
15
POMP: A Theoretical Approach to Mitigate Forgetting in Finetuning Multi-Modal Models
ICLR 2026Rejected
二作16
Beyond Hard Supervised Fine-tuning: Enhancing Image-text Alignment of Strong Models with Weak Models
ICLR 2026Rejected
二作24
Density-Aware Translation of Spurious Correlations in Zero-Shot VLMs
ICLR 2026Rejected
通讯11
Leveraging Dark Knowledge for Intrinsic Multimodal Out-of-Distribution Detection
ICLR 2026Rejected
二作24
Do Vision-Language Models Reason Like Humans? Exploring the Functional Roles of Attention Heads
ICLR 2026Rejected
23
Geometry-Guided Adversarial Prompt Detection via Curvature and Local Intrinsic Dimension
ICLR 2026Rejected
通讯14
GeoDetect: Geometric Adversarial Detection for VLPs
ICLR 2026Rejected
通讯25
What Do VLMs See? Benchmarking Vision-Language Models on Ambiguous Images
ICLR 2026Rejected
11
Characterising Universal Jailbreak Features and Refusal Direction in LLMs
ICLR 2026Withdrawn
通讯29
Cognitive Mirrors: Exploring the Diverse Functional Roles of Attention Heads in LLM Reasoning
NeurIPS 2025Poster
9
X-Transfer Attacks: Towards Super Transferable Adversarial Attacks on CLIP
ICML 2025Poster
二作14
Multi-level Certified Defense Against Poisoning Attacks in Offline Reinforcement Learning
ICLR 2025Poster
17
Fortifying Time Series: DTW-Certified Robust Anomaly Detection
NeurIPS 2025Poster
通讯20
Detecting Backdoor Samples in Contrastive Language Image Pretraining
ICLR 2025Poster
二作24
Attention-based Graph Coreset Labeling for Active Learning
ICLR 2025Rejected
三作28
HOGT: High-Order Graph Transformers
ICLR 2025Rejected
32
CURVALID: A Geometrically-guided Adversarial Prompt Detection
ICLR 2025Rejected
三作33
TUAP: Targeted Universal Adversarial Perturbations for CLIP
ICLR 2025Rejected
二作15
GAD-VLP: Geometric Adversarial Detection for Vision-Language Pre-Trained Models
ICLR 2025Rejected
三作13
Exploring Weak-to-Strong Generalization for CLIP-based Classification
ICLR 2025Withdrawn
二作