影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
98.65/100
前 0.1%
全站排名 #31
发表论文54 篇
平均评分
年均产出18.0 篇/年
Martin Vechev
研究方向
safe and secure AI · AI for mathematics · LLM security · AI for programming languages and program analysis · AI and symbolic reasoning · combining logic and AI · trustworthy AI
14
The Open Proof Corpus: A Large-Scale Study of LLM-Generated Mathematical Proofs
ICLR 2026Poster
通讯15
Watch your steps: Dormant Adversarial Behaviors that Activate upon LLM Finetuning
ICLR 2026Oral
通讯13
LLM Fingerprinting via Semantically Conditioned Watermarks
ICLR 2026Oral
通讯17
Fewer Weights, More Problems: A Practical Attack on LLM Pruning
ICLR 2026Poster
通讯17
Constrained Decoding of Diffusion LLMs with Context-Free Grammars
ICLR 2026Poster
三作14
BrokenMath: A Benchmark for Sycophancy in Theorem Proving with LLMs
ICLR 2026Rejected
三作11
Dual Randomized Smoothing: Beyond Global Noise Variance
ICLR 2026Poster
三作25
Watermarking Diffusion Language Models
ICLR 2026Poster
通讯13
Expressiveness of Multi-Neuron Convex Relaxations in Neural Network Certification
ICLR 2026Poster
三作11
ToolFuzz: Automated Agent Tool Testing
ICLR 2026Rejected
通讯21
Pay Attention to the Triggers: Constructing Backdoors That Survive Distillation
ICLR 2026Rejected
通讯17
Adaptive Generation of Bias-Eliciting Questions for LLMs
ICLR 2026Rejected
通讯15
AutoBaxBench: Bootstrapping Code Security Benchmarking
ICLR 2026Rejected
通讯16
Watermarking Autoregressive Image Generation
NeurIPS 2025Poster
19
MixAT: Combining Continuous and Discrete Adversarial Training for LLMs
NeurIPS 2025Poster
通讯22
Black-Box Detection of Language Model Watermarks
ICLR 2025Poster
通讯9
MathConstruct: Challenging LLM Reasoning with Constructive Proofs
ICML 2025Poster
通讯17
Polyrating: A Cost-Effective and Bias-Aware Rating System for LLM Evaluation
ICLR 2025Poster
三作19
Ward: Provable RAG Dataset Inference via LLM Watermarks
ICLR 2025Poster
通讯7
Discovering Spoofing Attempts on Language Model Watermarks
ICML 2025Poster
通讯11
A Unified Approach to Routing and Cascading for LLMs
ICML 2025Poster
三作13
Black-Box Adversarial Attacks on LLM-Based Code Completion
ICML 2025Poster
通讯13
Mind the Gap: A Practical Attack on GGUF Quantization
ICML 2025Poster
通讯23
Language Models are Advanced Anonymizers
ICLR 2025Poster
通讯32
Discovering Clues of Spoofed LM Watermarks
ICLR 2025Rejected
通讯35
GRAIN: Exact Graph Reconstruction from Gradients
ICLR 2025Poster
通讯13
Average Certified Radius is a Poor Metric for Randomized Smoothing
ICML 2025Poster
通讯25
Black-Box Adversarial Attacks on LLM-Based Code Completion
ICLR 2025Rejected
通讯24
A Unified Approach to Routing and Cascading for LLMs
ICLR 2025Rejected
三作11
Automated Benchmark Generation for Repository-Level Coding Tasks
ICML 2025Poster
三作9
BaxBench: Can LLMs Generate Correct and Secure Backends?
ICML 2025Spotlight
通讯22
CTBench: A Library and Benchmark for Certified Training
ICLR 2025Rejected
三作25
AlphaIntegrator: Transformer Action Search for Symbolic Integration Proofs
ICLR 2025Rejected
三作8
CTBench: A Library and Benchmark for Certified Training
ICML 2025Poster
三作13
Average Certified Radius is a Poor Metric for Randomized Smoothing
ICLR 2025Rejected
通讯14
Evading Data Contamination Detection for Language Models is (too) Easy
ICLR 2025Rejected
通讯5
Multi-Neuron Unleashes Expressivity of ReLU Networks Under Convex Relaxation
ICLR 2025Withdrawn
三作16
Gaussian Loss Smoothing Enables Certified Training with Tight Convex Relaxations
ICLR 2025Rejected
通讯