影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
60.66/100
前 4.7%
全站排名 #3,014
发表论文11 篇
平均评分
年均产出3.7 篇/年
Javier Rando
研究方向
AI Safety · Safety · Security · Robustness · Large Language Models · Red-Teaming
8
AutoAdvExBench: Benchmarking Autonomous Exploitation of Adversarial Example Defenses
ICML 2025Oral
三作20
Adversarial Perturbations Cannot Reliably Protect Artists From Generative AI
ICLR 2025Spotlight
二作33
Measuring Non-Adversarial Reproduction of Training Data in Large Language Models
ICLR 2025Poster
二作25
Scalable Extraction of Training Data from Aligned, Production Language Models
ICLR 2025Poster
二作21
AutoAdvExBench: Benchmarking Autonomous Exploitation of Adversarial Example Defenses
ICLR 2025Rejected
三作28
Persistent Pre-training Poisoning of LLMs
ICLR 2025Poster
二作26
Gradient-based Jailbreak Images for Multimodal Fusion Models
ICLR 2025Rejected
一作