影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
65.85/100
前 3.4%
全站排名 #2,199
发表论文13 篇
平均评分
年均产出6.5 篇/年
21
Any-Depth Alignment: Unlocking Innate Safety Alignment of LLMs to Any-Depth
ICLR 2026Poster
一作46
ARMs: Adaptive Red-Teaming Agent against Multimodal Models with Plug-and-Play Attacks
ICLR 2026Poster
17
GraphQ-LM: Scalable Graph Representation for Large Language Models via Residual Vector Quantization
ICLR 2026Rejected
一作24
MMDT: Decoding the Trustworthiness and Safety of Multimodal Foundation Models
ICLR 2025Poster
二作42
EIA: ENVIRONMENTAL INJECTION ATTACK ON GENERALIST WEB AGENTS FOR PRIVACY LEAKAGE
ICLR 2025Poster
12
AdvAgent: Controllable Blackbox Red-teaming on Web Agents
ICML 2025Poster
三作10
SafeAuto: Knowledge-Enhanced Safe Autonomous Driving with Multimodal Foundation Models
ICML 2025Poster
一作11
GuardAgent: Safeguard LLM Agents via Knowledge-Enabled Reasoning
ICML 2025Poster
29
GuardAgent: Safeguard LLM Agent by a Guard Agent via Knowledge-Enabled Reasoning
ICLR 2025Rejected
20
KnowHalu: Multi-Form Knowledge Enhanced Hallucination Detection
ICLR 2025Rejected
一作11
UDora: A Unified Red Teaming Framework against LLM Agents by Dynamically Hijacking Their Own Reasoning
ICML 2025Poster
一作23
AdvWeb: Controllable Black-box Attacks on VLM-powered Web Agents
ICLR 2025Rejected
三作23
SafeAuto: Knowledge-Enhanced Safe Autonomous Driving with Multimodal Foundation Models
ICLR 2025Rejected
一作