影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
79.14/100
前 1.5%
全站排名 #948
发表论文26 篇
平均评分
年均产出8.7 篇/年
Xiangyu Zhang
研究方向
large language models · large vision-language models · large multimodal models · representative learning · self-supervised learning · deep learning · neural networks · CNNs · computer vision · recognition · object detection
21
Mixture-of-Experts Can Surpass Dense LLMs Under Strictly Equal Resource
ICLR 2026Oral
12
NoiseAR: AutoRegressing Initial Noise Prior for Diffusion Models
ICLR 2026Rejected
三作20
MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation
ICLR 2026Poster
37
NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale
ICLR 2026Oral
22
Predictable Scale (Part II) --- Farseer: A Refined Scaling Law in LLMs
NeurIPS 2025Spotlight
25
Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model
NeurIPS 2025Poster
23
GUI Exploration Lab: Enhancing Screen Navigation in Agents via Multi-Turn Reinforcement Learning
NeurIPS 2025Poster
20
Unhackable Temporal Reward for Scalable Video MLLMs
ICLR 2025Poster
22
Perception-R1: Pioneering Perception Policy with Reinforcement Learning
NeurIPS 2025Poster
34
Open Vision Reasoner: Transferring Linguistic Cognitive Behavior for Visual Reasoning
NeurIPS 2025Poster
8
Perception in Reflection
ICML 2025Poster
32
DreamBench++: A Human-Aligned Benchmark for Personalized Image Generation
ICLR 2025Poster
31
Vision Foundation Models as Effective Visual Tokenizers for Autoregressive Generation
NeurIPS 2025Poster
30
Reconstructive Visual Instruction Tuning
ICLR 2025Poster
37
Glad: A Streaming Scene Generator for Autonomous Driving
ICLR 2025Poster
通讯21
PerPO: Perceptual Preference Optimization via Discriminative Rewarding
ICLR 2025Rejected
通讯5
Is a 3D-Tokenized LLM the Key to Reliable Autonomous Driving?
ICLR 2025Withdrawn
通讯5
Continual LLaVA: Continual Instruction Tuning in Large Vision-Language Models
ICLR 2025Withdrawn
11
General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model
ICLR 2025Withdrawn
通讯11
Continuous Semi-Implicit Models
ICML 2025Poster
23
Generalizable Dynamic Radiance Field in Egocentric View
ICLR 2025Rejected
25
Image Generation with Channel-wise Quantization
ICLR 2025Rejected
二作5
MOTRv3: Release-Fetch Supervision for End-to-End Multi-Object Tracking
ICLR 2025Withdrawn
5
PADriver: Towards Personalized Autonomous Driving
ICLR 2025Withdrawn
通讯