影响力指数
79.14/100
前 1.5%
全站排名 #948
发表论文26
平均评分5.5
年均产出8.7 篇/年

Xiangyu Zhang

Principal Researcher@MEGVII Technology·中国·OpenReview
研究方向

large language models · large vision-language models · large multimodal models · representative learning · self-supervised learning · deep learning · neural networks · CNNs · computer vision · recognition · object detection

8.2
22

Predictable Scale (Part II) --- Farseer: A Refined Scaling Law in LLMs

NeurIPS 2025Spotlight
7.8
25

Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model

NeurIPS 2025Poster
6.8
23

GUI Exploration Lab: Enhancing Screen Navigation in Agents via Multi-Turn Reinforcement Learning

NeurIPS 2025Poster
6.5
20

Unhackable Temporal Reward for Scalable Video MLLMs

ICLR 2025Poster
6.4
22

Perception-R1: Pioneering Perception Policy with Reinforcement Learning

NeurIPS 2025Poster
6.4
34

Open Vision Reasoner: Transferring Linguistic Cognitive Behavior for Visual Reasoning

NeurIPS 2025Poster
6.3
8

Perception in Reflection

ICML 2025Poster
6.0
32

DreamBench++: A Human-Aligned Benchmark for Personalized Image Generation

ICLR 2025Poster
6.0
31

Vision Foundation Models as Effective Visual Tokenizers for Autoregressive Generation

NeurIPS 2025Poster
5.8
30

Reconstructive Visual Instruction Tuning

ICLR 2025Poster
5.6
37

Glad: A Streaming Scene Generator for Autonomous Driving

ICLR 2025Poster
通讯
5.0
21

PerPO: Perceptual Preference Optimization via Discriminative Rewarding

ICLR 2025Rejected
通讯
5.0
5

Is a 3D-Tokenized LLM the Key to Reliable Autonomous Driving?

ICLR 2025Withdrawn
通讯
4.8
5

Continual LLaVA: Continual Instruction Tuning in Large Vision-Language Models

ICLR 2025Withdrawn
4.8
11

General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model

ICLR 2025Withdrawn
通讯
4.4
11

Continuous Semi-Implicit Models

ICML 2025Poster
4.2
23

Generalizable Dynamic Radiance Field in Egocentric View

ICLR 2025Rejected
4.0
25

Image Generation with Channel-wise Quantization

ICLR 2025Rejected
二作
4.0
5

MOTRv3: Release-Fetch Supervision for End-to-End Multi-Object Tracking

ICLR 2025Withdrawn
4.0
5

PADriver: Towards Personalized Autonomous Driving

ICLR 2025Withdrawn
通讯