影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
56.39/100
前 5.9%
全站排名 #3,819
发表论文16 篇
平均评分
年均产出8.0 篇/年
26
AudioTrust: Benchmarking The Multifaceted Trustworthiness of Audio Large Language Models
ICLR 2026Poster
通讯18
Tug-of-War No More: Harmonizing Accuracy and Robustness in Vision-Language Models via Stability-Aware Task Vector Merging
ICLR 2026Poster
24
OrthAlign: Orthogonal Subspace Decomposition for Non-Interfering Multi-Objective Alignment
ICLR 2026Poster
27
Learning to Rank Chain-of-Thought: Using a Small Model
ICLR 2026Desk Rejected
14
HAI-Eval: Measuring Human-AI Synergy in Collaborative Coding
ICLR 2026Desk Rejected
24
A-MemGuard: A Proactive Defense Framework for LLM-Based Agent Memory
ICLR 2026Rejected
34
MedSentry: Understanding and Mitigating Safety Risks in Medical LLM Multi-Agent Systems
ICLR 2026Rejected
22
Can LLMs Refuse Questions They Do Not Know? Measuring Knowledge-Aware Refusal in Factual Tasks
ICLR 2026Poster
21
GTD: Dynamic Generation of Multi LLM Agents Communication Topologies with Graph Diffusion Models
ICLR 2026Withdrawn
4
Energy-Driven Steering: Reducing False Refusals in Large Language Models
ICLR 2026Withdrawn
通讯5
UniErase: Towards Balanced and Precise Unlearning in Language Models
ICLR 2026Withdrawn
32
GuardReasoner-VL: Safeguarding VLMs via Reinforced Reasoning
NeurIPS 2025Poster
26
Adversarial Attacks against Closed-Source MLLMs via Feature Optimal Alignment
NeurIPS 2025Poster
19
AgentAuditor: Human-level Safety and Security Evaluation for LLM Agents
NeurIPS 2025Poster
25
SeCon-RAG: A Two-Stage Semantic Filtering and Conflict-Free Framework for Trustworthy RAG
NeurIPS 2025Poster
18
Prompt-Independent Safe Decoding to Restrain Unsafe Image Generation for Text-to-Image Models against White-Box Adversary
ICLR 2025Rejected