影响力指数
84.67/100
前 1%
全站排名 #618
发表论文31
平均评分4.8
年均产出10.3 篇/年

Xue Liu

Full Professor@Mohamed bin Zayed University of Artificial Intelligence·阿联酋·OpenReview
研究方向

Recommender Systems · Reinforcement Learning · Machine Learning

6.7
15

Escaping Policy Contraction: Contraction-Aware PPO (CaPPO) for Stable Language Model Fine-Tuning

ICLR 2026Poster
三作
6.0
16

Spatial CAPTCHA: Generatively Benchmarking Spatial Reasoning for Human-Machine Differentiation

ICLR 2026Poster
通讯
5.5
25

Pedagogically-Inspired Data Synthesis for Language Model Knowledge Distillation

ICLR 2026Poster
5.5
29

Adversarial Encoding Perturbation and Synthesis for Set Representation Auxiliary Learning

ICLR 2026Poster
通讯
5.5
18

Tequila: Trapping-free Ternary Quantization for Large Language Models

ICLR 2026Poster
5.3
17

RAEE: A Robust Retrieval-Augmented Early Exit Framework for Efficient Inference

ICLR 2026Poster
4.7
16

Prima.cpp: Fast 30-70B LLM Inference on Heterogeneous and Low-Resource Home Clusters

ICLR 2026Poster
通讯
4.5
18

Which Heads Matter for Reasoning? RL-Guided KV Cache Compression

ICLR 2026Rejected
4.5
19

Beyond Semantic Similarity: Reducing Unnecessary API Calls via Behavior-Aligned Retriever

ICLR 2026Rejected
4.5
21

Exploring the Trade-off between Quality and Diversity of Language Models during Reinforcement Learning

ICLR 2026Rejected
通讯
4.5
33

RECODE-H: A Benchmark for Research Code Development with Interactive Human Feedback

ICLR 2026Poster
3.5
5

Mixture of Experts Characteristic Function Embeddings for Heterogeneous Fraud Graphs

ICLR 2026Withdrawn
三作
3.3
8

Optimizing User Profiles via Contextual Bandits for Retrieval-Augmented LLM Personalization

ICLR 2026Withdrawn
3.2
6

Light-Search: Reducing Retrieval Cost in RAG via Curriculum-Based Policy Optimization

ICLR 2026Withdrawn
3.0
12

On Why Form Shapes Reasoning: Structuring Latent Program Networks with Category-Theoretic Constraints

ICLR 2026Rejected
二作
3.0
6

Reasoning at the Right Length: Adaptive Budget Forcing for Efficient and Accurate LLM Inference

ICLR 2026Withdrawn
三作
3.0
16

Diffusion Large Language Models for Black-Box Optimization

ICLR 2026Rejected
通讯
2.5
7

Recall-First Moderation via Distribution-Preserving Augmentation and Committee-Diverse Retrieval

ICLR 2026Rejected
二作