影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
85.18/100
前 0.9%
全站排名 #593
发表论文36 篇
平均评分
年均产出12.0 篇/年
20
AdAEM: An Adaptively and Automated Extensible Measurement of LLMs' Value Difference
ICLR 2026Oral
通讯19
Harnessing Temporal Databases for Systematic Evaluation of Factual Time-Sensitive Question-Answering in LLMs
ICLR 2026Poster
三作22
CAReDiO: Enhancing Cultural Alignment of LLM via Representativeness and Distinctiveness Guided Data Optimization
ICLR 2026Rejected
通讯21
PICACO: Pluralistic In-Context Value Alignment via Total Correlation Optimization
ICLR 2026Rejected
通讯18
To Think or Not To Think, That is The Question for LLM Reasoning in Theory of Mind Tasks
ICLR 2026Rejected
通讯23
Social-R1: Enhancing Social Intelligence in LLMs through Human-like Reinforced Reasoning
ICLR 2026Rejected
30
MMA-ASIA: A Multilingual and Multimodal Alignment Framework for Culturally Grounded Evaluation
ICLR 2026Withdrawn
5
Eliciting Human-like Social Reasoning in Large Language Models
ICLR 2026Withdrawn
9
From Assistants to Companions: Towards the Usefulness of Improving Theory of Mind for Human-AI Symbiosis
ICLR 2026Withdrawn
通讯9
Beyond BFI: The CSI for Enhanced Reliability and Validity in Evaluating LLM Personality Traits
ICLR 2026Rejected
27
Counterfactual Reasoning for Steerable Pluralistic Value Alignment of Large Language Models
NeurIPS 2025Poster
通讯17
Raising the Bar: Investigating the Values of Large Language Models via Generative Evolving Testing
ICML 2025Poster
通讯26
Unveiling the Learning Mind of Language Models: A Cognitive Framework and Empirical Study
NeurIPS 2025Poster
40
Raising the Bar: Investigating the Values of Large Language Models via Generative Evolving Testing
ICLR 2025Rejected
通讯19
Why Do You Answer Like That? Psychological Analysis on Underlying Connections between LLM's Values and Safety Risks
ICLR 2025Rejected
14
Leveraging Implicit Sentiments: Enhancing Reliability and Validity in Psychological Trait Evaluation of LLMs
ICLR 2025Rejected
16
Elephant in the Room: Unveiling the Pitfalls of Human Proxies in Alignment
ICLR 2025Rejected
6
Can I Understand What I Create? Self-Knowledge Evaluation of Large Language Models
ICLR 2025Rejected