影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
83.11/100
前 1.1%
全站排名 #714
发表论文46 篇
平均评分
年均产出15.3 篇/年
Jiaheng Liu
研究方向
Large Language Models · Model Acceleration · Point Cloud Understanding · Face Recognition
19
Tricks or Traps? A Deep Dive into RL for LLM Reasoning
ICLR 2026Poster
25
ScaleLong: A Multi-Timescale Benchmark for Long Video Understanding
ICLR 2026Poster
23
IV-Bench: A Benchmark for Image-Grounded Video Perception and Reasoning in Multimodal LLMs
ICLR 2026Poster
31
IF-VidCap: Can Video Caption Models Follow Instructions?
ICLR 2026Poster
通讯11
YuE: Scaling Open Foundation Models for Long-Form Music Generation
ICLR 2026Poster
23
Flash-Searcher: Fast and Effective Web Agents via DAG-Based Parallel Execution
ICLR 2026Poster
33
ArtifactsBench: Bridging the Visual-Interactive Gap in LLM Code Generation Evaluation
ICLR 2026Rejected
33
VideoSearch Reasoner: Boosting Multimodal Reward Models through Think with Image Reasoning
ICLR 2026Rejected
通讯19
TaskCraft: Automated Generation of Agentic Tasks
ICLR 2026Poster
30
Inverse IFEval: Can LLMs Unlearn Stubborn Training Conventions to Follow Real Instructions?
ICLR 2026Poster
29
DESIGNER: Design-Logic-Guided Multidisciplinary Data Synthesis for LLM Reasoning
ICLR 2026Poster
18
Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL
ICLR 2026Rejected
31
SafeDialBench: A Fine-Grained Safety Evaluation Benchmark for Large Language Models in Multi-Turn Dialogues with Diverse Jailbreak Attacks
ICLR 2026Poster
17
ACADREASON: Exploring the Limits of Reasoning Models with Academic Research Problems
ICLR 2026Poster
26
HiPO: Hybrid Policy Optimization for Dynamic Reasoning in LLMs
ICLR 2026Withdrawn
通讯44
OmniVideoBench: Towards Audio-Visual Understanding Evaluation for Omni MLLMs
ICLR 2026Poster
通讯17
MM-BrowseComp: A Comprehensive Benchmark for Multimodal Browsing Agents
ICLR 2026Withdrawn
46
Reconstructing KV Caches with Cross-Layer Fusion for Enhanced Transformers
ICLR 2026Poster
5
SPEAR: Structured Pruning for Spiking Neural Networks via Synaptic Operation Estimation and Reinforcement Learning
ICLR 2026Withdrawn
15
Beyond Safe Answers: A Benchmark for Evaluating True Risk Awareness in Large Reasoning Models
ICLR 2026Rejected
5
Agent KB: Leveraging Cross-Domain Experience for Agentic Problem Solving
ICLR 2026Withdrawn
5
MT-Video-Bench: A Holistic Video Understanding Benchmark for Evaluating Multimodal LLMs in Multi-Turn Dialogues
ICLR 2026Withdrawn
通讯5
LVCap-Eval: Towards Holistic Long Video Caption Evaluation for Multimodal LLMs
ICLR 2026Withdrawn
5
USB: A Comprehensive and Unified Safety Evaluation Benchmark for Multimodal Large Language Models
ICLR 2026Withdrawn
13
DGPO: Mitigating Likelihood Displacement with Bidirectional KL Divergence Gap
ICLR 2026Rejected
19
KORGym: A Dynamic Game Platform for LLM Reasoning Evaluation
NeurIPS 2025Spotlight
三作33
Flow-GRPO: Training Flow Matching Models via Online RL
NeurIPS 2025Poster
29
KOR-Bench: Benchmarking Language Models on Knowledge-Orthogonal Reasoning Tasks
ICLR 2025Poster
22
McEval: Massively Multilingual Code Evaluation
ICLR 2025Poster
14
MuPT: A Generative Symbolic Music Pretrained Transformer
ICLR 2025Poster
21
Towards Visualization-of-Thought Jailbreak Attack against Large Visual Language Models
NeurIPS 2025Poster
26
LIME: LESS IS MORE FOR MLLM EVALUATION
ICLR 2025Rejected
28
MTU-Bench: A Multi-granularity Tool-Use Benchmark for Large Language Models
ICLR 2025Poster
19
OmniBench: Towards The Future of Universal Omni-Language Models
ICLR 2025Rejected
37
M2rc-Eval: Massively Multilingual Repository-level Code Completion Evaluation
ICLR 2025Rejected
一作28
MIO: A Foundation Model on Multimodal Tokens
ICLR 2025Rejected
32
AutoKaggle: A Multi-Agent Framework for Autonomous Data Science Competitions
ICLR 2025Rejected
20
HelloBench: Evaluating Long Text Generation Capabilities of Large Language Models
ICLR 2025Withdrawn
6
ING-VP: MLLMs Cannot Play Easy Vision-based Games Yet
ICLR 2025Rejected
6
CodeChain: An Open, Million-scale Dataset for Code Language Models at the Repository Level
ICLR 2025Withdrawn
6
Can MLLMs Understand the Deep Implication Behind Chinese Images?
ICLR 2025Withdrawn