影响力指数
79.95/100
前 1.4%
全站排名 #899
发表论文35
平均评分4.9
年均产出17.5 篇/年

Xing Sun

Principal Researcher@Tencent YouTu Lab·中国·OpenReview
研究方向

LLM & MLLM · Person ReID

6.5
17

Attend to the Active: Structure-Aware Dynamic Attention in LLMs for Compositional Instruction Following

ICLR 2026Poster
6.0
19

PoLi-RL: A Point-to-List Reinforcement Learning Framework for Conditional Semantic Textual Similarity

ICLR 2026Poster
5.5
24

Process-Level Trajectory Evaluation for Environment Configuration in Software Engineering Agents

ICLR 2026Poster
5.5
19

VITA-E: A Dual-Model Framework for Real-Time, Interruptible, and Concurrent Human-Robot Interaction

ICLR 2026Rejected
5.0
30

Count Counts: Motivating Exploration in LLM Reasoning with Count-based Intrinsic Rewards

ICLR 2026Poster
5.0
21

GraphRAG-Bench: Challenging Domain-specific Reasoning for Evaluating Graph Retrieval-Augmented Generation

ICLR 2026Rejected
4.7
9

FactGuard: Detecting Unanswerable Questions in Long-Context Texts for Reliable LLM Responses

ICLR 2026Rejected
通讯
4.7
18

RAR: Reversing Visual Attention Re-Sinking for Unlocking Potential in Multimodal Large Language Models

ICLR 2026Poster
4.5
17

When should I search more: Adaptive Complex Query Optimization with Reinforcement Learning

ICLR 2026Withdrawn
通讯
4.5
21

ASPD: Unlocking Adaptive Serial-Parallel Decoding by Exploring Intrinsic Parallelism in LLMs

ICLR 2026Rejected
通讯
4.5
18

DeepOmni: Towards Seamless and Smart Speech Interaction with Adaptive Modality-Specific MoE

ICLR 2026Rejected
通讯
4.5
20

CUARewardBench: Benchmark for Evaluating Reward Models on Computer-using Agent Trajectories

ICLR 2026Rejected
通讯
4.0
16

HiChunk: Evaluating and Enhancing Retrieval-Augmented Generation with Hierarchical Chunking

ICLR 2026Withdrawn
通讯
4.0
24

Youtu-GraphRAG: Vertically Unified Agents for Graph Retrieval-Augmented Complex Reasoning

ICLR 2026Poster
通讯
4.0
5

VITA-VLA: Efficiently Teaching Vision-Language Models to Act via Action Expert Distillation

ICLR 2026Withdrawn
4.0
21

Learn the Ropes, Then Trust the Wins: Self-imitation with Progressive Exploration for Agentic Reinforcement Learning

ICLR 2026Poster
通讯
3.5
15

Improve LLM Pre-training with RL-Guided Annealing

ICLR 2026Rejected
3.5
26

CoDiEmb: A Collaborative yet Distinct Framework for Unified Representation Learning in Information Retrieval and Semantic Textual Similarity

ICLR 2026Rejected
通讯
3.5
5

TACO: Think-Answer Consistency for Optimized Long-Chain Reasoning and Efficient Data Learning via Reinforcement Learning in LVLMs

ICLR 2026Withdrawn
3.5
15

Training-Free Group Relative Policy Optimization

ICLR 2026Rejected
通讯
3.3
11

APTBench: Benchmarking Agentic Potential of Base LLMs During Pre-Training

ICLR 2026Rejected
通讯
3.3
5

RoRecomp: Enhancing Reasoning Efficiency via Rollout Response Recomposition in Reinforcement Learning

ICLR 2026Rejected
2.0
5

Dimensional Debiasing via Multi-Agent Correction

ICLR 2026Withdrawn