影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
92.43/100
前 0.4%
全站排名 #265
发表论文41 篇
平均评分
年均产出13.7 篇/年
Xuanjing Huang
研究方向
large language model · text mining · fundamental NLP
22
AgentGym-RL: An Open-Source Framework to Train LLM Agents for Long-Horizon Decision Making via Multi-Turn RL
ICLR 2026Oral
14
RECAST: Expanding the Boundaries of LLMs' Complex Instruction Following with Multi-Constraint Data
ICLR 2026Poster
22
R-Horizon: How Far Can Your Large Reasoning Model Really Go in Breadth and Depth?
ICLR 2026Poster
31
Biologically Plausible Learning via Bidirectional Spike-Based Distillation
ICLR 2026Poster
22
Game-RL: Synthesizing Multimodal Verifiable Game Data to Boost VLMs' General Reasoning
ICLR 2026Poster
通讯26
EgoNight: Towards Egocentric Vision Understanding at Night with a Challenging Benchmark
ICLR 2026Poster
19
Critique-RL: Training Language Models For Critiquing Through Two-Stage Reinforcement Learning
ICLR 2026Poster
通讯28
BAPO: Stabilizing Off-Policy Reinforcement Learning for LLMs via Balanced Policy Optimization with Adaptive Clipping
ICLR 2026Poster
通讯17
Why Reinforcement Fine-Tuning Enables MLLMs Preserve Prior Knowledge Better: A Data Perspective
ICLR 2026Poster
20
Adaptive Curriculum Strategies: Stabilizing Reinforcement Learning for Large Language Models
ICLR 2026Rejected
21
Unlocking the Essence of Beauty: Advanced Aesthetic Reasoning with Relative-Absolute Policy Optimization
ICLR 2026Poster
通讯20
Two Minds Better Than One: Collaborative Reward Modeling for LLM Alignment
ICLR 2026Rejected
24
WorldPM: Understanding Scaling Patterns in Human Preference Modeling
ICLR 2026Rejected
18
Search or Think? Rethinking Iterative RAG from An Entropy Perspective
ICLR 2026Withdrawn
26
Mitigating Position Bias in Transformers via Layer-Specific Positional Embedding Scaling
ICLR 2026Withdrawn
11
MDAR: A Multi-scene Dynamic Audio Reasoning Benchmark
ICLR 2026Withdrawn
通讯13
From Scores to Preferences: Redefining MOS Benchmarking for Speech Quality Reward Modeling
ICLR 2026Withdrawn
通讯5
Context-Aware Alignment: Adapting Large Language Models to Individual Historical Data
ICLR 2026Withdrawn
6
Model Utility Law: Evaluating LLMs beyond Performance via Mechanistically Interpretable Metric
ICLR 2026Rejected
23
EvaLearn: Quantifying the Learning Capability and Efficiency of LLMs via Sequential Problem Solving
NeurIPS 2025Poster
通讯28
Understanding Parametric and Contextual Knowledge Reconciliation within Large Language Models
NeurIPS 2025Spotlight
通讯19
Toward Relative Positional Encoding in Spiking Transformers
NeurIPS 2025Spotlight
24
Pre-Trained Policy Discriminators are General Reward Models
NeurIPS 2025Poster
14
Have the VLMs Lost Confidence? A Study of Sycophancy in VLMs
ICLR 2025Poster
通讯24
Domain-RAG: Retrieval-Guided Compositional Image Generation for Cross-Domain Few-Shot Object Detection
NeurIPS 2025Poster
19
Effective Length Extrapolation via Dimension-Wise Positional Embeddings Manipulation
COLM 2025Poster
通讯26
Distill Visual Chart Reasoning Ability from LLMs to MLLMs
ICLR 2025Rejected
通讯22
RMB: Comprehensively benchmarking reward models in LLM alignment
ICLR 2025Poster
通讯20
How Jailbreak Defenses Work and Ensemble? A Mechanistic Investigation
ICLR 2025Rejected
44
AgentGym: Evaluating and Evolving Large Language Model-based Agents across Diverse Envronments
ICLR 2025Rejected
4
SPA-VL: A Comprehensive Safety Preference Alignment Dataset for Vision Language Model
ICLR 2025Withdrawn
45
Tell Me What You Don't Know: Enhancing Refusal Capabilities of Role-Playing Agents via Representation Space Analysis and Editing
ICLR 2025Rejected
通讯15
Dendritic Localized Learning: Toward Biologically Plausible Algorithm
ICML 2025Poster
通讯15
Progressive LLM Alignments Using Two-Player Games
ICLR 2025Rejected