影响力指数
94.64/100
前 0.3%
全站排名 #184
发表论文46
平均评分5.3
年均产出15.3 篇/年

Junyang Lin

Principal Researcher@Alibaba Group·中国·OpenReview
研究方向

Large Language Models · Multimodal Pretraining · Natural Language Processing

7.0
13

Language Confusion Gate: Language-Aware Decoding Through Model Self-Distillation

ICLR 2026Poster
通讯
6.5
19

VideoAgentTrek: Computer-Use Pretraining from Unlabeled Videos

ICLR 2026Poster
5.5
15

PlotCraft: Pushing the Limits of LLMs for Complex and Interactive Data Visualization

ICLR 2026Withdrawn
通讯
5.5
22

Omni-Captioner: Data Pipeline, Models, and Benchmark for Omni Detailed Perception

ICLR 2026Poster
5.5
30

From Narrow to Panoramic Vision: Attention-Guided Cold-Start Reshapes Multimodal Reasoning

ICLR 2026Poster
5.5
14

MegaFlow: Large-Scale Distributed Orchestration System for the Agentic Era

ICLR 2026Rejected
5.3
26

A$^2$Search: Ambiguity-Aware Question Answering with Reinforcement Learning

ICLR 2026Poster
通讯
5.0
13

Revisiting Multimodal Positional Encoding in Vision–Language Models

ICLR 2026Poster
5.0
16

Overcoming Joint Intractability with Lossless Hierarchical Speculative Decoding

ICLR 2026Oral
4.5
20

RefineX: Learning to Refine Pre-training Data at Scale from Expert-Guided Programs

ICLR 2026Desk Rejected
4.5
18

SWE-RM: Execution-free Feedback for Software Engineering Agents

ICLR 2026Poster
4.5
32

Winning the Pruning Gamble: A Unified Approach to Joint Sample and Token Pruning for Efficient Supervised Fine-Tuning

ICLR 2026Desk Rejected
4.5
28

TriSpec: Ternary Speculative Decoding via Lightweight Proxy Verification

ICLR 2026Rejected
通讯
4.5
21

WavReward: Spoken Dialogue Models With Generalist Reward Evaluators

ICLR 2026Rejected
4.5
24

WorldPM: Understanding Scaling Patterns in Human Preference Modeling

ICLR 2026Rejected
通讯
4.4
14

C$^3$-Bench: Evaluating and Achieving Controllable Code Completion in Code LLM

ICLR 2026Withdrawn
通讯
4.0
11

One Model to Critique Them All: Rewarding Agentic Tool-Use via Efficient Reasoning

ICLR 2026Withdrawn
4.0
13

Understanding DeepResearch via Reports

ICLR 2026Rejected
4.0
21

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback

ICLR 2026Rejected
通讯
4.0
20

Beyond Turn Limits: Training Deep Search Agents with Dynamic Context Window

ICLR 2026Rejected
通讯
3.5
6

ReAlign: Safety-Aligning Reasoning Models with Verifier-Guided Reinforcement Learning

ICLR 2026Rejected
三作
2.5
5

Reinforcement Learning for Symbolic Graphics Code with Visual Feedback

ICLR 2026Withdrawn
8.7
21

Gated Attention for Large Language Models: Non-linearity, Sparsity, and Attention-Sink-Free

NeurIPS 2025Oral
通讯
7.8
26

Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning

NeurIPS 2025Poster
通讯
7.3
26

Chain of Execution Supervision Promotes General Reasoning in Large Language Models

NeurIPS 2025Poster
7.3
19

Parallel Scaling Law for Language Models

NeurIPS 2025Poster
7.0
16

OpenHands: An Open Platform for AI Software Developers as Generalist Agents

ICLR 2025Poster
6.6
9

MARGE: Improving Math Reasoning with Guided Exploration

ICML 2025Poster
6.4
27

Teaching Language Models to Reason with Tools

NeurIPS 2025Poster
6.3
20

Self-Evolving Critique Abilities in Large Language Models

COLM 2025Poster
通讯
6.3
11

Efficient Long Context Fine-tuning with Chunk Flow

ICML 2025Poster
6.2
18

A Spark of Vision-Language Intelligence: 2-Dimensional Autoregressive Transformer for Efficient Finegrained Image Generation

ICLR 2025Poster
6.1
15

CateKV: On Sequential Consistency for Long-Context LLM Inference Acceleration

ICML 2025Poster
6.1
9

Synthesizing Software Engineering Data in a Test-Driven Manner

ICML 2025Poster
通讯
6.0
38

DataMan: Data Manager for Pre-training Large Language Models

ICLR 2025Poster
6.0
26

CARE: Decoding-Time Safety Alignment via Rollback and Introspection Intervention

NeurIPS 2025Poster
5.8
23

Rotated Runtime Smooth: Training-Free Activation Smoother for accurate INT4 inference

ICLR 2025Poster
4.7
21

Disentangling Reasoning Tokens and Boilerplate Tokens For Language Model Fine-tuning

ICLR 2025Withdrawn
4.6
42

Analyzing and Mitigating Inconsistency in Discrete Audio Tokens for Neural Codec Language Models

ICLR 2025Rejected
通讯
4.4
23

Rethinking Data Selection at Scale: Random Selection is Almost All You Need

ICLR 2025Rejected
通讯
3.8
6

Language Models can Self-Lengthen to Generate Long Texts

ICLR 2025Withdrawn
通讯
3.5
5

Aligning Large Language Models via Self-Steering Optimization

ICLR 2025Withdrawn
通讯