影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
86.32/100
前 0.8%
全站排名 #535
发表论文19 篇
平均评分
年均产出9.5 篇/年
27
Scaling Embedding Layers in Language Models
NeurIPS 2025Poster
19
Quantifying Cross-Modality Memorization in Vision-Language Models
NeurIPS 2025Poster
二作21
Crosslingual Capabilities and Knowledge Barriers in Multilingual Large Language Models
COLM 2025Poster
三作24
SORRY-Bench: Systematically Evaluating Large Language Model Safety Refusal
ICLR 2025Poster
30
On Evaluating the Durability of Safeguards for Open-Weight LLMs
ICLR 2025Poster
20
GMValuator: Similarity-based Data Valuation for Generative Models
ICLR 2025Poster
18
Unlearn and Burn: Adversarial Machine Unlearning Requests Destroy Model Accuracy
ICLR 2025Poster
一作7
Exploring and Mitigating Adversarial Manipulation of Voting-Based Leaderboards
ICML 2025Oral
一作23
MUSE: Machine Unlearning Six-Way Evaluation for Language Models
ICLR 2025Poster
三作10
MATH-Perturb: Benchmarking LLMs' Math Reasoning Abilities against Hard Perturbations
ICML 2025Poster
23
Crosslingual Capabilities and Knowledge Barriers in Multilingual Large Language Models
ICLR 2025Rejected
三作10
Scaling Laws for Differentially Private Language Models
ICML 2025Poster
二作13
Scaling Embedding Layers in Language Models
ICML 2025Rejected
31
On Memorization of Large Language Models in Logical Reasoning
ICLR 2025Rejected
二作17
Fantastic Copyrighted Beasts and How (Not) to Generate Them
ICLR 2025Poster
二作18
Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation
ICLR 2024Spotlight
一作13
Mind the Privacy Unit! User-Level Differential Privacy for Language Model Fine-Tuning
COLM 2024Poster
三作17
Detecting Pretraining Data from Large Language Models
ICLR 2024Poster
19
LabelDP-Pro: Learning with Label Differential Privacy via Projections
ICLR 2024Poster
二作