影响力指数
83.11/100
前 1.1%
全站排名 #714
发表论文46
平均评分5.2
年均产出15.3 篇/年

Jiaheng Liu

Assistant Professor@Nanjing University·中国·OpenReview
研究方向

Large Language Models · Model Acceleration · Point Cloud Understanding · Face Recognition

6.0
19

Tricks or Traps? A Deep Dive into RL for LLM Reasoning

ICLR 2026Poster
6.0
25

ScaleLong: A Multi-Timescale Benchmark for Long Video Understanding

ICLR 2026Poster
6.0
23

IV-Bench: A Benchmark for Image-Grounded Video Perception and Reasoning in Multimodal LLMs

ICLR 2026Poster
6.0
31

IF-VidCap: Can Video Caption Models Follow Instructions?

ICLR 2026Poster
通讯
6.0
11

YuE: Scaling Open Foundation Models for Long-Form Music Generation

ICLR 2026Poster
5.6
23

Flash-Searcher: Fast and Effective Web Agents via DAG-Based Parallel Execution

ICLR 2026Poster
5.5
33

ArtifactsBench: Bridging the Visual-Interactive Gap in LLM Code Generation Evaluation

ICLR 2026Rejected
5.5
33

VideoSearch Reasoner: Boosting Multimodal Reward Models through Think with Image Reasoning

ICLR 2026Rejected
通讯
5.5
19

TaskCraft: Automated Generation of Agentic Tasks

ICLR 2026Poster
5.3
30

Inverse IFEval: Can LLMs Unlearn Stubborn Training Conventions to Follow Real Instructions?

ICLR 2026Poster
4.8
29

DESIGNER: Design-Logic-Guided Multidisciplinary Data Synthesis for LLM Reasoning

ICLR 2026Poster
4.7
18

Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL

ICLR 2026Rejected
4.5
31

SafeDialBench: A Fine-Grained Safety Evaluation Benchmark for Large Language Models in Multi-Turn Dialogues with Diverse Jailbreak Attacks

ICLR 2026Poster
4.5
17

ACADREASON: Exploring the Limits of Reasoning Models with Academic Research Problems

ICLR 2026Poster
4.5
26

HiPO: Hybrid Policy Optimization for Dynamic Reasoning in LLMs

ICLR 2026Withdrawn
通讯
4.5
44

OmniVideoBench: Towards Audio-Visual Understanding Evaluation for Omni MLLMs

ICLR 2026Poster
通讯
4.0
17

MM-BrowseComp: A Comprehensive Benchmark for Multimodal Browsing Agents

ICLR 2026Withdrawn
4.0
46

Reconstructing KV Caches with Cross-Layer Fusion for Enhanced Transformers

ICLR 2026Poster
4.0
5

SPEAR: Structured Pruning for Spiking Neural Networks via Synaptic Operation Estimation and Reinforcement Learning

ICLR 2026Withdrawn
3.5
15

Beyond Safe Answers: A Benchmark for Evaluating True Risk Awareness in Large Reasoning Models

ICLR 2026Rejected
3.5
5

Agent KB: Leveraging Cross-Domain Experience for Agentic Problem Solving

ICLR 2026Withdrawn
3.5
5

MT-Video-Bench: A Holistic Video Understanding Benchmark for Evaluating Multimodal LLMs in Multi-Turn Dialogues

ICLR 2026Withdrawn
通讯
3.5
5

LVCap-Eval: Towards Holistic Long Video Caption Evaluation for Multimodal LLMs

ICLR 2026Withdrawn
3.0
5

USB: A Comprehensive and Unified Safety Evaluation Benchmark for Multimodal Large Language Models

ICLR 2026Withdrawn
2.5
13

DGPO: Mitigating Likelihood Displacement with Bidirectional KL Divergence Gap

ICLR 2026Rejected