影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
74.04/100
前 2%
全站排名 #1,303
发表论文38 篇
平均评分
年均产出12.7 篇/年
Tong Yu
研究方向
Vision Language Models · Natural Language Processing · Reinforcement Learning · Recommender System · Data Intelligence
15
MASS-DPO: Multi-negative Active Sample Selection for Direct Policy Optimization
ICLR 2026Rejected
20
Importance Sampling for Multi-Negative Multimodal Direct Preference Optimization
ICLR 2026Poster
15
AMPS: Adaptive Modality Preference Steering via Functional Entropy
ICLR 2026Rejected
18
Precise Attribute Intensity Control in Large Language Models via Targeted Representation Editing
ICLR 2026Desk Rejected
14
VisR-Bench: An Empirical Study on Visual Retrieval-Augmented Generation for Multilingual Long Document Understanding
ICLR 2026Rejected
25
Mitigating Forgetting Between Supervised and Reinforcement Learning Yields Stronger Reasoners
ICLR 2026Rejected
三作12
FERA: Uncertainty-aware Federated Reasoning for Large Language Models
ICLR 2026Rejected
6
WaterFlow: Fast and Robust Watermarking in the Latent Fourier Domain
ICLR 2026Rejected
15
A Personalized Conversational Benchmark: Towards Simulating Personalized Conversations
ICLR 2026Withdrawn
4
WS-GRPO: Weakly-Supervised Group-Relative Policy Optimization
ICLR 2026Withdrawn
6
Beyond Static Retrieval Policies: Task-Aware Adaptive RAG With METAR
ICLR 2026Rejected
二作5
Contextual Bandits with LLM-Derived Priors and Adaptive Calibration
ICLR 2026Withdrawn
16
In-context Ranking Preference Optimization
COLM 2025Poster
23
Listwise Preference Diffusion Optimization for User Behavior Trajectories Prediction
NeurIPS 2025Poster
35
OCEAN: Offline Chain-of-thought Evaluation and Alignment in Large Language Models
ICLR 2025Poster
16
Traceable and Explainable Multimodal Large Language Models: An Information-Theoretic View
COLM 2025Poster
15
SV-RAG: LoRA-Contextualizing Adaptation of MLLMs for Long Document Understanding
ICLR 2025Poster
13
A Survey on Personalized and Pluralistic Preference Alignment in Large Language Models
COLM 2025Poster
9
Federated In-Context Learning: Iterative Refinement for Improved Answer Quality
ICML 2025Poster
14
VipAct: Visual-Perception Enhancement via Specialized VLM Agent Collaboration and Tool-use
ICLR 2025Rejected
三作15
CodeLutra: Boosting LLM Code Generation via Preference-Guided Refinement
ICLR 2025Withdrawn
三作11
Offline-to-Online Reinforcement Learning with Classifier-Free Diffusion Generation
ICML 2025Poster
13
Fully Dynamic Embedding into $\ell_p$ Spaces
ICML 2025Poster
通讯21
CoMMIT: Coordinated Instruction Tuning for Multimodal Large Language Models
ICLR 2025Rejected
三作5
FigCaps-HF: A Figure-to-Caption Generative Framework and Benchmark with Human Feedback
ICLR 2025Withdrawn
5
Offline-to-Online Reinforcement Learning with Classifier-Free Diffusion Generation
ICLR 2025Rejected