影响力指数
98.68/100
前 0.1%
全站排名 #27
发表论文66
平均评分5.4
年均产出22.0 篇/年

Furu Wei

Distinguished Scientist@Microsoft Research·中国·OpenReview
研究方向

general ai · foundation model · deep learning · natural language procesing

6.7
26

VibeVoice: Expressive Podcast Generation with Next-Token Diffusion

ICLR 2026Oral
通讯
6.5
17

Synergizing Understanding and Generation with Interleaved Analyzing-Drafting Thinking

ICLR 2026Poster
5.5
28

VisCodex: Unified Multimodal Code Generation via Merging Vision and Coding Models

ICLR 2026Poster
通讯
5.5
23

From Abstract to Contextual: What LLMs Still Cannot Do in Mathematics

ICLR 2026Poster
通讯
5.0
27

Code Aesthetics with Agentic Reward Feedback

ICLR 2026Poster
通讯
5.0
17

Multimodal Latent Language Modeling with Next-Token Diffusion

ICLR 2026Rejected
通讯
5.0
16

11Plus-Bench: Demystifying Multimodal LLM Spatial Reasoning with Cognitive-Inspired Analysis

ICLR 2026Rejected
通讯
5.0
18

Learning To Draft: Adaptive Speculative Decoding with Reinforcement Learning

ICLR 2026Poster
5.0
15

Geometric-Mean Policy Optimization

ICLR 2026Poster
通讯
4.8
25

AlignDiff: Exploiting Model-Intrinsic Information for Better Preference Data Selection

ICLR 2026Rejected
4.7
10

BitNet Distillation

ICLR 2026Rejected
通讯
4.5
29

Scaling Laws for Fully Sparsely-Activated Large Language Models

ICLR 2026Rejected
通讯
4.5
18

QueST: Incentivizing LLMs to Generate Difficult Problems

ICLR 2026Rejected
通讯
4.5
36

Breaking Training Bottlenecks: Effective Reinforcement Learning for Modern Coding Models

ICLR 2026Rejected
通讯
4.5
14

Rectified Sparse Attention for Efficient Long-Sequence Generation

ICLR 2026Withdrawn
通讯
4.5
15

WildLong: Synthesizing Realistic Long-Context Instruction Data at Scale

ICLR 2026Rejected
通讯
4.5
16

Two Pathways to Truthfulness: On the Intrinsic Encoding of LLM Hallucinations

ICLR 2026Withdrawn
4.0
5

Towards Stable and Effective Reinforcement learning for Mixture-of-Experts

ICLR 2026Withdrawn
通讯
4.0
17

DocReward: A Document Reward Model for Structuring and Stylizing

ICLR 2026Withdrawn
通讯
3.5
12

Thinking Augmented Pre-training

ICLR 2026Rejected
通讯
3.5
14

Learning to Refine: Self-Refinement of Parallel Reasoning in LLMs

ICLR 2026Withdrawn
3.5
14

Latent Sketchpad: Sketching Visual Thoughts to Elicit Multimodal Reasoning in MLLMs

ICLR 2026Rejected
通讯
2.5
5

Information-Preserving Reformulation of Reasoning Traces for Antidistillation

ICLR 2026Withdrawn
通讯
2.5
5

On-Policy RL with Optimal Reward Baseline

ICLR 2026Withdrawn
通讯
2.5
5

Efficient RL Training for Reasoning Models via Length-Aware Optimization

ICLR 2026Withdrawn
8.0
25

Data Selection via Optimal Control for Language Models

ICLR 2025Oral
8.0
21

Differential Transformer

ICLR 2025Oral
通讯
7.8
23

Think Only When You Need with Large Hybrid-Reasoning Models

NeurIPS 2025Poster
通讯
7.3
17

Preference Optimization for Reasoning with Pseudo Feedback

ICLR 2025Spotlight
通讯
7.0
25

Semi-Parametric Retrieval via Binary Bag-of-Tokens Index

ICLR 2025Poster
三作
7.0
17

Generative Representational Instruction Tuning

ICLR 2025Poster
6.8
17

Chain-of-Retrieval Augmented Generation

NeurIPS 2025Poster
通讯
6.6
31

Self-Boosting Large Language Models with Synthetic Preference Data

ICLR 2025Poster
通讯
6.5
27

Scaling Laws of Synthetic Data for Language Model

COLM 2025Poster
通讯
6.4
29

Towards Thinking-Optimal Scaling of Test-Time Compute for LLM Reasoning

NeurIPS 2025Poster
通讯
6.4
29

Reward Reasoning Models

NeurIPS 2025Poster
通讯
6.3
31

ARLON: Boosting Diffusion Transformers with Autoregressive Models for Long Video Generation

ICLR 2025Poster
通讯
6.1
11

Imagine While Reasoning in Space: Multimodal Visualization-of-Thought

ICML 2025Poster
通讯
6.0
17

Scaling Optimal LR Across Token Horizons

ICLR 2025Poster
5.8
16

E5-V: Universal Embeddings with Multimodal Large Language Models

ICLR 2025Rejected
5.3
23

Synthetic Data (Almost) from Scratch: Generalized Instruction Tuning for Language Models

ICLR 2025Rejected
通讯
5.0
6

VALL-E 2: Neural Codec Language Models are Human Parity Zero-Shot Text to Speech Synthesizers

ICLR 2025Rejected
通讯
4.8
14

Q-Sparse: All Large Language Models can be Fully Sparsely-Activated

ICLR 2025Rejected
通讯
4.8
5

MoRA: High-Rank Updating for Parameter-Efficient Fine-Tuning

ICLR 2025Withdrawn
4.7
23

Textual Aesthetics in Large Language Models

ICLR 2025Withdrawn
通讯
4.3
21

One Language, Many Gaps: Evaluating Dialect Fairness and Robustness of Large Language Models in Reasoning Tasks

ICLR 2025Withdrawn
通讯
3.5
19

Next Block Prediction: Video Generation via Semi-Auto-Regressive Modeling

ICLR 2025Rejected
通讯