影响力指数
92.43/100
前 0.4%
全站排名 #265
发表论文41
平均评分5.4
年均产出13.7 篇/年

Xuanjing Huang

Full Professor@Fudan University·中国·OpenReview
研究方向

large language model · text mining · fundamental NLP

7.0
22

AgentGym-RL: An Open-Source Framework to Train LLM Agents for Long-Horizon Decision Making via Multi-Turn RL

ICLR 2026Oral
6.7
14

RECAST: Expanding the Boundaries of LLMs' Complex Instruction Following with Multi-Constraint Data

ICLR 2026Poster
6.0
22

R-Horizon: How Far Can Your Large Reasoning Model Really Go in Breadth and Depth?

ICLR 2026Poster
5.6
31

Biologically Plausible Learning via Bidirectional Spike-Based Distillation

ICLR 2026Poster
5.2
22

Game-RL: Synthesizing Multimodal Verifiable Game Data to Boost VLMs' General Reasoning

ICLR 2026Poster
通讯
5.0
26

EgoNight: Towards Egocentric Vision Understanding at Night with a Challenging Benchmark

ICLR 2026Poster
5.0
19

Critique-RL: Training Language Models For Critiquing Through Two-Stage Reinforcement Learning

ICLR 2026Poster
通讯
5.0
28

BAPO: Stabilizing Off-Policy Reinforcement Learning for LLMs via Balanced Policy Optimization with Adaptive Clipping

ICLR 2026Poster
通讯
4.7
17

Why Reinforcement Fine-Tuning Enables MLLMs Preserve Prior Knowledge Better: A Data Perspective

ICLR 2026Poster
4.7
20

Adaptive Curriculum Strategies: Stabilizing Reinforcement Learning for Large Language Models

ICLR 2026Rejected
4.5
21

Unlocking the Essence of Beauty: Advanced Aesthetic Reasoning with Relative-Absolute Policy Optimization

ICLR 2026Poster
通讯
4.5
20

Two Minds Better Than One: Collaborative Reward Modeling for LLM Alignment

ICLR 2026Rejected
4.5
24

WorldPM: Understanding Scaling Patterns in Human Preference Modeling

ICLR 2026Rejected
4.0
18

Search or Think? Rethinking Iterative RAG from An Entropy Perspective

ICLR 2026Withdrawn
4.0
26

Mitigating Position Bias in Transformers via Layer-Specific Positional Embedding Scaling

ICLR 2026Withdrawn
3.5
11

MDAR: A Multi-scene Dynamic Audio Reasoning Benchmark

ICLR 2026Withdrawn
通讯
3.3
13

From Scores to Preferences: Redefining MOS Benchmarking for Speech Quality Reward Modeling

ICLR 2026Withdrawn
通讯
2.5
5

Context-Aware Alignment: Adapting Large Language Models to Individual Historical Data

ICLR 2026Withdrawn
2.5
6

Model Utility Law: Evaluating LLMs beyond Performance via Mechanistically Interpretable Metric

ICLR 2026Rejected
8.2
23

EvaLearn: Quantifying the Learning Capability and Efficiency of LLMs via Sequential Problem Solving

NeurIPS 2025Poster
通讯
7.8
28

Understanding Parametric and Contextual Knowledge Reconciliation within Large Language Models

NeurIPS 2025Spotlight
通讯
7.3
19

Toward Relative Positional Encoding in Spiking Transformers

NeurIPS 2025Spotlight
7.3
24

Pre-Trained Policy Discriminators are General Reward Models

NeurIPS 2025Poster
6.7
14

Have the VLMs Lost Confidence? A Study of Sycophancy in VLMs

ICLR 2025Poster
通讯
6.4
24

Domain-RAG: Retrieval-Guided Compositional Image Generation for Cross-Domain Few-Shot Object Detection

NeurIPS 2025Poster
6.3
19

Effective Length Extrapolation via Dimension-Wise Positional Embeddings Manipulation

COLM 2025Poster
通讯
6.0
26

Distill Visual Chart Reasoning Ability from LLMs to MLLMs

ICLR 2025Rejected
通讯
6.0
22

RMB: Comprehensively benchmarking reward models in LLM alignment

ICLR 2025Poster
通讯
5.8
20

How Jailbreak Defenses Work and Ensemble? A Mechanistic Investigation

ICLR 2025Rejected
5.8
44

AgentGym: Evaluating and Evolving Large Language Model-based Agents across Diverse Envronments

ICLR 2025Rejected
5.3
4

SPA-VL: A Comprehensive Safety Preference Alignment Dataset for Vision Language Model

ICLR 2025Withdrawn
5.2
45

Tell Me What You Don't Know: Enhancing Refusal Capabilities of Role-Playing Agents via Representation Space Analysis and Editing

ICLR 2025Rejected
通讯
4.6
15

Dendritic Localized Learning: Toward Biologically Plausible Algorithm

ICML 2025Poster
通讯
4.5
15

Progressive LLM Alignments Using Two-Player Games

ICLR 2025Rejected