影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
4.56/100
超过 14.2%
全站排名 #55,257
发表论文9 篇
平均评分
年均产出4.5 篇/年
Jose Hernandez-Orallo
研究方向
Large language models · data science · evaluation metrics · machine learning · AI evaluation · philosophy of AI
16
11Plus-Bench: Demystifying Multimodal LLM Spatial Reasoning with Cognitive-Inspired Analysis
ICLR 2026Rejected
20
Inferring Capabilities from Task Performance with Bayesian Triangulation
ICLR 2026Rejected
通讯5
Psychometric Personality Shaping Modulates Capabilities and Safety in Language Models
ICLR 2026Withdrawn
通讯11
From Human-Level AI Tales to AI Levelling Human Scales
ICLR 2026Rejected
通讯21
Personalized Safety in LLMs: A Benchmark and A Planning-Based Agent Approach
NeurIPS 2025Poster
12
Leaving the barn door open for Clever Hans: Simple features predict LLM benchmark answers
ICLR 2025Rejected
通讯12
What should an AI assessor optimise for?
ICLR 2025Rejected
二作20
100 instances is all you need: predicting LLM success by testing on a few instances
ICLR 2025Rejected
三作14
Relative Drawing Identification Complexity is Invariant to Modality in Vision-Language Models
ICLR 2025Rejected
通讯