影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
66.66/100
前 3.2%
全站排名 #2,091
发表论文27 篇
平均评分
年均产出9.0 篇/年
Kui Ren
研究方向
Data security and privacy protection · AI security · Cyber security
10
Mitigating Over-Refusal in Adversarial Tuning via Subspace-guided Sample Selection
ICLR 2026Rejected
通讯18
Towards Real-world Debiasing: Rethinking Evaluation, Challenge, and Solution
ICLR 2026Rejected
通讯11
Untargeted Jailbreak Attack
ICLR 2026Withdrawn
通讯16
Explainable Token-level Noise Filtering for LLM Fine-tuning Datasets
ICLR 2026Poster
通讯12
The Robustness-Security Paradox: Channel-Aware Feature Learning for Adversarial Watermark Exploitation
ICLR 2026Rejected
通讯16
Dynamic Target Attack
ICLR 2026Rejected
通讯5
Module-Aware Parameter-Efficient Machine Unlearning on Transformers
ICLR 2026Withdrawn
通讯5
CoKV: Optimizing LLM Inference with Game-Theoretic Adaptive KV Cache
ICLR 2026Withdrawn
通讯4
CARP: Causal Alignment of Reward Models via Response-to-Prompt Prediction
ICLR 2026Withdrawn
通讯8
Rethinking Shapley Value for Data Contribution
ICLR 2026Withdrawn
通讯5
Towards Evaluation for Real-World LLM Unlearning
ICLR 2026Withdrawn
通讯5
CollabMask: Explainable Neuron Collaboration with Gradient Masks for LLM Fine-Tuning
ICLR 2026Withdrawn
通讯6
ZeroSecBench: Fine-grained and Robust Evaluation for Secure Code Generation
ICLR 2026Rejected
25
WMCopier: Forging Invisible Watermarks on Arbitrary Images
NeurIPS 2025Poster
通讯17
Textual Unlearning Gives a False Sense of Unlearning
ICML 2025Poster
通讯25
Taught Well Learned Ill: Towards Distillation-conditional Backdoor Attack
NeurIPS 2025Poster
通讯29
Robust Representation Consistency Model via Contrastive Denoising
ICLR 2025Poster
19
JudgeRail: Harnessing Open-Source LLMs for Fast Harmful Text Detection with Judicial Prompting and Logit Rectification
ICLR 2025Rejected
通讯23
Towards Real World Debiasing: A Fine-grained Analysis On Spurious Correlation
ICLR 2025Rejected
通讯49
REFINE: Inversion-Free Backdoor Defense via Model Reprogramming
ICLR 2025Poster
通讯36
DAG-SHAP: Feature Attribution in DAG based on Edge Intervention
ICLR 2025Withdrawn
通讯28
Structure-Aware Parameter-Efficient Machine Unlearning on Transformer Models
ICLR 2025Rejected
通讯6
Prompt-Consistency Image Generation (PCIG): A Unified Framework Integrating LLMs, Knowledge Graphs, and Controllable Diffusion Models
ICLR 2025Withdrawn
通讯36
Towards False-claim-resistant Model Ownership Verification via Targeted Fingerprint
ICLR 2025Rejected
通讯7
A Comprehensive Deepfake Detector Assessment Platform
ICLR 2025Rejected
通讯