影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
44.35/100
前 11.4%
全站排名 #7,339
发表论文14 篇
平均评分
年均产出4.7 篇/年
Arthur Conmy
16
Thought Anchors: Which LLM Reasoning Steps Matter?
ICLR 2026Rejected
通讯16
Chain-of-Thought Reasoning In The Wild Is Not Always Faithful
ICLR 2026Rejected
通讯6
Base Models Know How to Reason, Thinking Models Learn When
ICLR 2026Withdrawn
11
Eliciting Secret Knowledge from Language Models
ICLR 2026Rejected
8
Fluid Reasoning Representations
ICLR 2026Rejected
三作10
SAEBench: A Comprehensive Benchmark for Sparse Autoencoders in Language Model Interpretability
ICML 2025Poster
22
Applying Sparse Autoencoders to Unlearn Knowledge in Language Models
ICLR 2025Rejected
三作11
Interpreting Attention Layer Outputs with Sparse Autoencoders
ICLR 2025Rejected
24
Scaling Sparse Feature Circuits For Studying In-Context Learning
ICLR 2025Rejected
通讯19
Jumping Ahead: Improving Reconstruction Fidelity with JumpReLU Sparse Autoencoders
ICLR 2025Rejected
11
Scaling Sparse Feature Circuits For Studying In-Context Learning
ICML 2025Poster
三作