影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
73.06/100
前 2.2%
全站排名 #1,389
发表论文19 篇
平均评分
年均产出6.3 篇/年
Zhihang Yuan
研究方向
Hardware and software co-optimization · Neural Network Quantization · Acceleration of Deep Learning · Efficient AI Algorithm
25
PCDVQ: Enhancing Vector Quantization for Large Language Models via Polar Coordinate Decoupling
ICLR 2026Withdrawn
三作24
Bootstrapping Language and Numerical Feedback for Reinforcement Learning in LLMs
ICLR 2026Rejected
三作6
SplitMeanFlow: Interval Splitting Consistency in Few-Step Generative Modeling
ICLR 2026Rejected
三作20
BMAttn: Block-Aligned Mixed-Precision Attention Quantization for LLM Inference
ICLR 2026Rejected
-1
INT v.s. FP: A Comprehensive Study of Fine-Grained Low-bit Quantization Formats
ICLR 2026Withdrawn
21
R2R: Efficiently Navigating Divergent Reasoning Paths with Small-Large Model Token Routing
NeurIPS 2025Poster
24
SAFEx: Analyzing Vulnerabilities of MoE-Based LLMs via Stable Safety-critical Expert Identification
NeurIPS 2025Poster
23
ASVD: Activation-aware Singular Value Decomposition for Compressing Large Language Models
ICLR 2025Rejected
一作21
MambaQuant: Quantizing the Mamba Family with Variance Aligned Rotation Methods
ICLR 2025Poster
33
OSTQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
ICLR 2025Poster
15
MoEQuant: Enhancing Quantization for Mixture-of-Experts Large Language Models via Expert-Balanced Sampling and Affinity Guidance
ICML 2025Poster
26
MQuant: Unleashing the Inference Potential of Multimodal Large Language Models via Full Static Quantization
ICLR 2025Rejected
通讯11
RWKVQuant: Quantizing the RWKV Family with Proxy Guided Hybrid of Scalar and Vector Quantization
ICML 2025Poster
29
I-LLM: Efficient Integer-Only Inference for Fully-Quantized Low-Bit Large Language Models
ICLR 2025Rejected
13
MxMoE: Mixed-precision Quantization for MoE with Accuracy and Performance Co-Design
ICML 2025Poster
三作5
A Closer Look at Time Steps is Worthy of Triple Speed-Up for Diffusion Model Training
ICLR 2025Withdrawn