影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
60.14/100
前 4.8%
全站排名 #3,089
发表论文10 篇
平均评分
年均产出3.3 篇/年
6
Thought Branches: Interpreting LLM Reasoning Requires Resampling
ICLR 2026Poster
三作10
Emergent Misalignment is Easy, Narrow Misalignment is Hard
ICLR 2026Poster
三作16
Chain-of-Thought Reasoning In The Wild Is Not Always Faithful
ICLR 2026Rejected
14
Steering Out-of-Distribution Generalization with Concept Ablation Fine-Tuning
ICLR 2026Rejected
11
Eliciting Secret Knowledge from Language Models
ICLR 2026Rejected
18
Do I Know This Entity? Knowledge Awareness and Hallucinations in Language Models
ICLR 2025Oral
三作22
Dense SAE Latents Are Features, Not Bugs
NeurIPS 2025Poster
10
Are Sparse Autoencoders Useful? A Case Study in Sparse Probing
ICML 2025Poster
三作19
Jumping Ahead: Improving Reconstruction Fidelity with JumpReLU Sparse Autoencoders
ICLR 2025Rejected
一作