影响力指数
88.11/100
前 0.7%
全站排名 #455
发表论文42
平均评分5.0
年均产出14.0 篇/年

Xuming Hu

Assistant Professor@The Hong Kong University of Science and Technology (Guangzhou)·中国·OpenReview
研究方向

Natural Language Processing · Machine Learning

7.0
13

PMark: Towards Robust and Distortion-free Semantic-level Watermarking with Channel Constraints

ICLR 2026Poster
5.5
21

You only need 4 extra tokens: Synergistic Test-time Adaptation for LLMs

ICLR 2026Rejected
5.3
13

CATMark: A Context-Aware Thresholding Framework for Robust Cross-Task Watermarking in Large Language Models

ICLR 2026Rejected
通讯
5.0
19

SaFeR-VLM: Toward Safety-aware Fine-grained Reasoning in Multimodal Models

ICLR 2026Rejected
5.0
26

Investigating Redundancy in Multimodal Large Language Models with Multiple Vision Encoders

ICLR 2026Poster
4.8
39

KnowMT-Bench: Benchmarking Knowledge-Intensive Long-Form Question Answering In Multi-Turn Dialogues

ICLR 2026Withdrawn
通讯
4.5
18

DiffAdapt: Difficulty-Adaptive Reasoning for Token-Efficient LLM Inference

ICLR 2026Poster
二作
4.5
37

Rethinking LLM Evaluation: Can We Evaluate LLMs with 200× Less Data?

ICLR 2026Poster
4.5
22

Zero in on Faithful Anchors: High-Fidelity Visual Token Condensation for Multimodal Large Language Models

ICLR 2026Rejected
通讯
4.4
23

Rethinking GCNs for the Traveling Salesman Problem: Are We Encoding Effectively?

ICLR 2026Rejected
4.0
18

Accelerate Diffusion Transformers with Feature Momentum

ICLR 2026Rejected
4.0
12

DocPruner: A Storage-Efficient Framework for Multi-Vector Visual Document Retrieval via Adaptive Patch-Level Embedding Pruning

ICLR 2026Rejected
通讯
3.5
6

Understanding-in-Generation: Reinforcing Generative Capability of Unified Model via Infusing Understanding into Generation

ICLR 2026Rejected
通讯
3.5
13

Distilling the Thought, Watermarking the Answer: A Principle Semantic Guided Watermark for Reasoning Large Language Models

ICLR 2026Poster
通讯
3.5
5

Consistent3DGen: Bridging Stochastic Generation and Deterministic Reconstruction for Image-to-3D Diffusion Models

ICLR 2026Withdrawn
3.5
5

Learning Robust Anymodal Segmentor with Unimodal and Cross-modal Distillation

ICLR 2026Withdrawn
通讯
3.3
4

Spot the Critical Words: Text-Guided Visual Token Pruning for Efficient Large Vision-Language Model Inference

ICLR 2026Withdrawn
通讯
3.0
16

FlowKV: Enhancing Multi-Turn Conversational Coherence in LLMs via Isolated Key-Value Cache Management

ICLR 2026Withdrawn
三作
3.0
12

Can LLMs Maintain Fundamental Abilities under KV Cache Compression?

ICLR 2026Withdrawn
3.0
5

MOSS-ChatV: Reinforcement Learning with Process Reasoning Reward for Video Temporal Reasoning

ICLR 2026Withdrawn
通讯
7.5
15

Can Watermarked LLMs be Identified by Users via Crafted Prompts?

ICLR 2025Spotlight
通讯
7.3
27

Don't Just Chase “Highlighted Tokens” in MLLMs: Revisiting Visual Holistic Context Retention

NeurIPS 2025Poster
通讯
6.6
15

OneForecast: A Universal Framework for Global and Regional Weather Forecasting

ICML 2025Poster
6.4
21

LoTA-QAF: Lossless Ternary Adaptation for Quantization-Aware Fine-Tuning

NeurIPS 2025Poster
通讯
6.4
33

ChunkKV: Semantic-Preserving KV Cache Compression for Efficient Long-Context LLM Inference

NeurIPS 2025Poster
6.1
10

Look Twice Before You Answer: Memory-Space Visual Retracing for Hallucination Mitigation in Multimodal Large Language Models

ICML 2025Poster
通讯
6.0
7

Gnothi Seauton: Empowering Faithful Self-Interpretability in Black-Box Transformers

ICLR 2025Poster
5.8
6

ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection

ICLR 2025Rejected
5.5
41

Mitigating Modality Prior-Induced Hallucinations in Multimodal Large Language Models via Deciphering Attention Causality

ICLR 2025Poster
通讯
5.5
11

RealRAG: Retrieval-augmented Realistic Image Generation via Self-reflective Contrastive Learning

ICML 2025Poster
通讯
5.3
6

ChunkKV: Semantic-Preserving KV Cache Compression for Efficient Long-Context LLM Inference

ICLR 2025Rejected
5.3
32

EXPLORING RESPONSE UNCERTAINTY IN MLLMS: AN EMPIRICAL EVALUATION UNDER MISLEADING SCENARIOS

ICLR 2025Rejected
通讯
5.0
25

LAIA-SQL: Enhancing Natural Language to SQL Generation in Multi-Table QA via Task Decomposition and Keyword Extraction

ICLR 2025Withdrawn
三作
4.8
8

MINER: Mining the Underlying Pattern of Modality-Specific Neurons in Multimodal Large Language Models

ICLR 2025Rejected
通讯
4.8
31

Unlocking Speech Instruction Data Potential with Query Rewriting

ICLR 2025Rejected
二作
4.8
34

Look Twice Before You Answer: Memory-Space Visual Retracing for Hallucination Mitigation in Multimodal Large Language Models

ICLR 2025Rejected
通讯
4.6
9

MAD: Multi-Alignment MEG-to-Text Decoding

ICLR 2025Withdrawn
4.6
33

Interpretable Contrastive Monte Carlo Tree Search Reasoning

ICLR 2025Withdrawn
4.2
6

DRUPI: Dataset Reduction Using Privileged Information

ICLR 2025Withdrawn