影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
47.35/100
前 9.8%
全站排名 #6,304
发表论文7 篇
平均评分
年均产出3.5 篇/年
Lei Yu
研究方向
AI agent · reinforcement learning · interpretability · AI alignment · large language models
12
Hallucination Reduction with CASAL: Contrastive Activation Steering for Amortized Learning
ICLR 2026Poster
三作13
VL-JEPA: Joint Embedding Predictive Architecture for Vision-language
ICLR 2026Poster
16
Misaligned Roles, Misplaced Images: Structural Input Perturbations Expose Multimodal Alignment Blind Spots
ICLR 2026Poster
37
Emergence of a High-Dimensional Abstraction Phase in Language Transformers
ICLR 2025Poster
20
Robust LLM safeguarding via refusal feature adversarial training
ICLR 2025Poster
一作46
Intrinsic Evaluation of Unlearning Using Parametric Knowledge Traces
ICLR 2025Rejected
二作47
GEOMETRIC SIGNATURES OF COMPOSITIONALITY ACROSS A LANGUAGE MODEL’S LIFETIME
ICLR 2025Withdrawn
三作