影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
53.53/100
前 6.9%
全站排名 #4,457
发表论文6 篇
平均评分
年均产出2.0 篇/年
18
Sparse Feature Circuits: Discovering and Editing Interpretable Causal Graphs in Language Models
ICLR 2025Oral
二作27
NNsight and NDIF: Democratizing Access to Open-Weight Foundation Model Internals
ICLR 2025Poster
10
SAEBench: A Comprehensive Benchmark for Sparse Autoencoders in Language Model Interpretability
ICML 2025Poster
二作