影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
64.23/100
前 3.7%
全站排名 #2,390
发表论文26 篇
平均评分
年均产出8.7 篇/年
20
DeepMath-103K: A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning
ICLR 2026Poster
23
Castle-in-the-Air: Evaluating MLLM Visual Abilities on Human Cognitive Benchmarks
ICLR 2026Rejected
19
UI-Ins: Enhancing GUI Grounding with Multi-Perspective Instruction as Reasoning
ICLR 2026Poster
15
ComboBench: Can LLMs Manipulate Physical Devices to Play Virtual Reality Games?
ICLR 2026Rejected
36
Towards Evaluating Fake Reasoning Bias in Language Models
ICLR 2026Withdrawn
29
The PIMMUR Principles: Ensuring Validity in Collective Behavior of LLM Societies
ICLR 2026Rejected
9
DRQA: Dynamic Reasoning Quota Allocation for Controlling Overthinking in Reasoning Large Language Models
ICLR 2026Withdrawn
5
The Social Welfare Function Leaderboard: When LLM Agents Allocate Social Welfare
ICLR 2026Withdrawn
5
Curing "Miracle Steps'' in LLM Math Reasoning with Rubric Rewards
ICLR 2026Withdrawn
15
Benchmarking MLLM-based Web Understanding: Reasoning, Robustness and Safety
ICLR 2026Rejected
通讯10
The Hunger Game Debate: On the Emergence of Over-Competition in Multi-Agent Systems
ICLR 2026Withdrawn
18
Language Models Do Not Have Human-Like Working Memory
ICLR 2026Rejected
三作23
Time-R1: Post-Training Large Vision Language Model for Temporal Video Grounding
NeurIPS 2025Poster
28
Trust, But Verify: A Self-Verification Approach to Reinforcement Learning with Verifiable Rewards
NeurIPS 2025Poster
23
Two Experts Are All You Need for Steering Thinking: Reinforcing Cognitive Effort in MoE Reasoning Models Without Additional Training
NeurIPS 2025Poster
11
On the Resilience of LLM-Based Multi-Agent Collaboration with Faulty Agents
ICML 2025Poster
23
Competing Large Language Models in Multi-Agent Gaming Environments
ICLR 2025Poster
31
On the Resilience of Multi-Agent Systems with Malicious Agents
ICLR 2025Rejected
15
Chain-of-Jailbreak Attack for Image Generation Models via Editing Step by Step
ICLR 2025Rejected
一作6
Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training
ICLR 2025Withdrawn
三作4
Insight Over Sight? Exploring the Vision-Knowledge Conflicts in Multimodal LLMs
ICLR 2025Withdrawn
二作