影响力指数
95.07/100
前 0.3%
全站排名 #167
发表论文46
平均评分5.3
年均产出15.3 篇/年

Jingren Zhou

Researcher@Alibaba Group·中国·OpenReview
研究方向

Machine Learning · Database Systems · Distributed Systems

6.5
27

WebWeaver: Structuring Web-Scale Evidence with Dynamic Outlines for Open-Ended Deep Research

ICLR 2026Poster
通讯
6.0
17

Demystifying Deep Search: A Holistic Evaluation with Hint-free Multi-Hop Questions and Factorised Metrics

ICLR 2026Poster
通讯
6.0
39

Empowering Efficiency and Efficacy in WebAgent via Enabling Info-Rich Seeking

ICLR 2026Poster
6.0
24

WebShaper: Agentically Data Synthesizing via Information-Seeking Formalization

ICLR 2026Poster
通讯
5.5
21

Expanding the Capability Frontier of LLM Agents with ZPD-Guided Data Synthesis

ICLR 2026Poster
5.5
10

AgentFold: Long-Horizon Web Agents with Proactive Context Folding

ICLR 2026Poster
5.5
17

Beyond Magnitude: Leveraging Direction of RLVR Updates for LLM Reasoning

ICLR 2026Poster
通讯
5.5
23

WebWatcher: Breaking New Frontiers of Vision-Language Deep Research Agent

ICLR 2026Poster
通讯
5.3
15

MaskSearch: Towards Scalable Agentic Pre-Training for Search-Enhanced Reasoning

ICLR 2026Rejected
通讯
5.0
27

On-Policy RL Meets Off-Policy Experts: Harmonizing Supervised Fine-Tuning and Reinforcement Learning via Dynamic Weighting

ICLR 2026Poster
通讯
5.0
14

ZeroSearch: Incentivize the Search Capability of LLMs without Searching

ICLR 2026Desk Rejected
通讯
5.0
20

Repurposing Synthetic Data for Fine-grained Search Agent Supervision

ICLR 2026Poster
5.0
17

WebSailor-V2: Bridging the Chasm to Proprietary Agents via Synthetic Data and Scalable Reinforcement Learning

ICLR 2026Poster
通讯
4.7
16

Scaling Agents via Continual Pre-training

ICLR 2026Poster
通讯
4.5
19

Sparse but Critical: A Token-Level Analysis of Distributional Shifts in RLVR Fine-Tuning of LLMs

ICLR 2026Poster
通讯
4.5
24

WorldPM: Understanding Scaling Patterns in Human Preference Modeling

ICLR 2026Rejected
4.4
33

BOTS: A Unified Framework for Bayesian Online Task Selection in LLM Reinforcement Finetuning

ICLR 2026Poster
通讯
4.0
33

IterResearch: Rethinking Long-Horizon Agents with Interaction Scaling

ICLR 2026Poster
通讯
4.0
27

ReSum: Unlocking Long-Horizon Search Intelligence via Context Summarization

ICLR 2026Rejected
通讯
4.0
13

Towards General Agentic Intelligence via Environment Scaling

ICLR 2026Withdrawn
通讯
3.5
6

MixLLM: Selecting Large Language Models with High-Quality Results and Minimum Inference Cost for Multi-Stage Complex Tasks

ICLR 2026Rejected
通讯
-1

WebSailor: Navigating Super-human Reasoning for Web Agent

ICLR 2026Desk Rejected
通讯
8.7
21

Gated Attention for Large Language Models: Non-linearity, Sparsity, and Attention-Sink-Free

NeurIPS 2025Oral
7.2
26

Self-play with Execution Feedback: Improving Instruction-following Capabilities of Large Language Models

ICLR 2025Spotlight
通讯
6.8
22

WebDancer: Towards Autonomous Information Seeking Agency

NeurIPS 2025Poster
通讯
6.8
38

mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models

ICLR 2025Poster
通讯
6.4
27

Provable Scaling Laws for the Test-Time Compute of Large Language Models

NeurIPS 2025Poster
通讯
6.4
22

ACE: All-round Creator and Editor Following Instructions via Diffusion Transformer

ICLR 2025Poster
通讯
6.3
26

Benchmarking Multimodal Retrieval Augmented Generation with Dynamic VQA Dataset and Self-adaptive Planning Agent

ICLR 2025Poster
6.0
16

Data-Juicer Sandbox: A Feedback-Driven Suite for Multimodal Data-Model Co-development

ICML 2025Spotlight
通讯
5.8
23

Rotated Runtime Smooth: Training-Free Activation Smoother for accurate INT4 inference

ICLR 2025Poster
通讯
5.8
29

Data-Juicer Sandbox: A Comprehensive Suite for Multimodal Data-Model Co-development

ICLR 2025Rejected
通讯
4.8
5

Group Diffusion Transformers are Unsupervised Multitask Learners

ICLR 2025Withdrawn
通讯
3.8
6

Language Models can Self-Lengthen to Generate Long Texts

ICLR 2025Withdrawn
3.5
5

Aligning Large Language Models via Self-Steering Optimization

ICLR 2025Withdrawn
3.0
16

On the Design and Analysis of LLM-Based Algorithms

ICLR 2025Rejected
通讯
3.0
11

Very Large-Scale Multi-Agent Simulation with LLM-Powered Agents

ICLR 2025Rejected
通讯