影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
84.64/100
前 1%
全站排名 #621
发表论文21 篇
平均评分
年均产出7.0 篇/年
Florian Tramèr
研究方向
adversarial examples · security · privacy
14
Pitfalls in Evaluating Language Model Forecasters
ICLR 2026Poster
通讯14
The Attacker Moves Second: Stronger Adaptive Attacks Bypass Defenses Against LLM Jailbreaks and Prompt Injections
ICLR 2026Rejected
19
Modal Aphasia: Can Unified Multimodal Models Describe Images From Memory?
ICLR 2026Poster
通讯25
OptiFluence: Scalable and Principled Design of Privacy Canaries
ICLR 2026Rejected
11
Black-box Optimization of LLM Outputs by Asking for Directions
ICLR 2026Rejected
通讯13
Can Reasoning Models Obfuscate Reasoning? Stress-Testing Chain-of-Thought Monitorability
ICLR 2026Desk Rejected
9
The Jailbreak Tax: How Useful are Your Jailbreak Outputs?
ICML 2025Spotlight
通讯8
AutoAdvExBench: Benchmarking Autonomous Exploitation of Adversarial Example Defenses
ICML 2025Oral
通讯20
Adversarial Perturbations Cannot Reliably Protect Artists From Generative AI
ICLR 2025Spotlight
通讯14
Consistency Checks for Language Model Forecasters
ICLR 2025Oral
通讯33
Measuring Non-Adversarial Reproduction of Training Data in Large Language Models
ICLR 2025Poster
通讯22
Adversarial Search Engine Optimization for Large Language Models
ICLR 2025Poster
三作25
Scalable Extraction of Training Data from Aligned, Production Language Models
ICLR 2025Poster
7
Exploring and Mitigating Adversarial Manipulation of Voting-Based Leaderboards
ICML 2025Oral
21
AutoAdvExBench: Benchmarking Autonomous Exploitation of Adversarial Example Defenses
ICLR 2025Rejected
通讯28
Persistent Pre-training Poisoning of LLMs
ICLR 2025Poster
26
Gradient-based Jailbreak Images for Multimodal Fusion Models
ICLR 2025Rejected
通讯20
Blind Baselines Beat Membership Inference Attacks for Foundation Models
ICLR 2025Rejected
三作