影响力指数
92.87/100
前 0.4%
全站排名 #254
发表论文44
平均评分5.2
年均产出14.7 篇/年

Yang Liu

Professor@Tsinghua University·中国·OpenReview
研究方向

Medical AI · AI for Science · Natural Language Processing

7.0
12

Kimi-Dev: Agentless Training as Skill Prior for SWE-agents

ICLR 2026Poster
6.5
17

DrugTrail: Interpretable Drug Discovery via Structured Reasoning and Druggability‑Tailored Preference Optimization

ICLR 2026Poster
6.0
23

Unified Biomolecular Trajectory Generation via Pretrained Variational Bridge

ICLR 2026Poster
三作
5.2
32

Doctor-R1: Mastering Clinical Inquiry with Experiential Agentic Reinforcement Learning

ICLR 2026Poster
通讯
5.0
20

Enhancing Large Language Model Reasoning via Selective Critical Token Fine-Tuning

ICLR 2026Rejected
4.8
25

AlignDiff: Exploiting Model-Intrinsic Information for Better Preference Data Selection

ICLR 2026Rejected
4.5
21

GraphPlan: Graph-enhanced Planning via Thinking LLMs for Embodied Agents

ICLR 2026Rejected
4.5
19

Agent-Environment Alignment via Automated Interface Generation

ICLR 2026Rejected
通讯
4.5
25

Visual Abstract Thinking: Enhancing Multimodal Reasoning via Visual Abstraction

ICLR 2026Rejected
通讯
4.4
26

Writing-RL: Advancing Long-form Writing via Adaptive Curriculum Reinforcement Learning

ICLR 2026Withdrawn
通讯
4.0
11

The Dialogue That Heals: A Comprehensive Evaluation of Doctor Agent's Inquiry Capability

ICLR 2026Withdrawn
通讯
4.0
14

Transparent and Robust RAG: Adaptive-Reward Reinforcement Learning for Decision Traceability

ICLR 2026Withdrawn
通讯
4.0
11

Inference-Time Scaling for Generalist Reward Modeling

ICLR 2026Rejected
4.0
20

MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding

ICLR 2026Rejected
通讯
3.5
21

Unveiling Over-Memorization in Finetuning LLMs for Reasoning Tasks

ICLR 2026Withdrawn
3.5
8

Benchmarking Long-Term Memory with Continuous dialogue Lifelogs

ICLR 2026Withdrawn
3.3
5

HBDrug3D: A Dataset and Benchmark for AI-Driven Heterobifunctional Molecule Design

ICLR 2026Rejected
通讯
3.2
36

UR$^2$: Unify RAG and Reasoning through Reinforcement Learning

ICLR 2026Withdrawn
通讯
2.7
9

SPEC-RL: Accelerating On-Policy Reinforcement Learning via Speculative Rollouts

ICLR 2026Withdrawn
7.8
22

MOF-BFN: Metal-Organic Frameworks Structure Prediction via Bayesian Flow Networks

NeurIPS 2025Poster
通讯
7.8
25

Learning 3D Anisotropic Noise Distributions Improves Molecular Force Fields

NeurIPS 2025Poster
7.0
22

Force-Guided Bridge Matching for Full-Atom Time-Coarsened Dynamics of Peptides

ICLR 2025Rejected
三作
7.0
7

UniMoMo: Unified Generative Modeling of 3D Molecules for De Novo Binder Design

ICML 2025Poster
通讯
6.4
34

Latent Retrieval Augmented Generation of Cross-Domain Protein Binders

NeurIPS 2025Poster
通讯
6.4
25

Beyond the Surface: Enhancing LLM-as-a-Judge Alignment with Human via Internal Representations

NeurIPS 2025Poster
6.3
17

FormaRL: Enhancing Autoformalization with no Labeled Data

COLM 2025Poster
通讯
5.8
19

StreamingBench: Assessing the Gap for MLLMs to Achieve Streaming Video Understanding

ICLR 2025Rejected
5.5
13

UniSim: A Unified Simulator for Time-Coarsened Dynamics of Biomolecules

ICML 2025Poster
三作
5.5
34

Advancing Language Multi-Agent Learning with Credit Re-Assignment for Interactive Environment Generalization

COLM 2025Poster
通讯
5.5
9

Zero-Shot Cyclic Peptide Design via Composable Geometric Constraints

ICML 2025Poster
通讯
4.4
6

VidEgoThink: Assessing Egocentric Video Understanding Capabilities for Embodied AI

ICLR 2025Withdrawn
通讯
4.3
18

ActiView: Evaluating Active Perception Ability for Multimodal Large Language Models

ICLR 2025Withdrawn
通讯
3.7
5

Enabling Weak LLMs to Judge Response Reliability via Meta Ranking

ICLR 2025Rejected
通讯
3.0
6

Enhancing Multi-Agent Learning in Real-World Interactive Environments through Process Reward Decomposition

ICLR 2025Rejected
通讯