影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
52.88/100
前 7.2%
全站排名 #4,610
发表论文10 篇
平均评分
年均产出5.0 篇/年
Hassan Sajjad
研究方向
machine translation · natural language processing · interpretation and analysis of deep neural network models · generative models · NLP · explainable AI · safety alignment · model editing
21
Data-centric Prediction Explanation via Kernelized Stein Discrepancy
ICLR 2025Poster
二作7
Explaining the role of Intrinsic Dimensionality in Adversarial Training
ICML 2025Poster
15
Dependency Parsing is More Parameter-Efficient with Normalization
NeurIPS 2025Poster
三作26
Resolving Lexical Bias in Edit Scoping with Projector Editor Networks
ICLR 2025Rejected
通讯15
Resolving Lexical Bias in Model Editing
ICML 2025Poster
通讯16
Towards Understanding the Feasibility of Machine Unlearning
ICLR 2025Rejected
二作21
SMAAT: Scalable Manifold-Aware Adversarial Training for Large Language Models
ICLR 2024Rejected
36
Representation Noising: A Defence Mechanism Against Harmful Finetuning
NeurIPS 2024Poster
10
Latent Concept-based Explanation of NLP Models
ICLR 2024Rejected
通讯9
Gazelle: A Multimodal Learning System Robust to Missing Modalities
ICLR 2024Withdrawn