影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
66.08/100
前 3.4%
全站排名 #2,171
发表论文18 篇
平均评分
年均产出9.0 篇/年
17
NLI : Non-uniform Linear Interpolation Approximation of Nonlinear Operations for Efficient LLMs Inference
ICLR 2026Poster
通讯34
KBVQ-MoE: KLT-guided SVD with Bias-Corrected Vector Quantization for MoE Large Language Models
ICLR 2026Poster
通讯32
SAES-SVD: Self-Adaptive Suppression of Accumulated and Local Errors for SVD-based LLM Compression
ICLR 2026Poster
二作25
PCDVQ: Enhancing Vector Quantization for Large Language Models via Polar Coordinate Decoupling
ICLR 2026Withdrawn
18
CMQuant: A Quantization-Aware Parameter-Efficient Fine-Tuning Framework for 4-Bit Consistency Models
ICLR 2026Rejected
二作13
DLLMQuant: A Post-Training Quantization Framework Tailored for Diffusion-Based Large Language Models
ICLR 2026Rejected
二作4
MSAVQ: Multi-dimensional Sensitivity-Aware Vector Quantization for Ultra-Low-Bit Vision-Language Models
ICLR 2026Desk Rejected
通讯39
BiMoE:Pushing the Limit of Post-Training Quantization for MoE-based LLMs
ICLR 2026Desk Rejected
通讯41
Towards W2A4 LLM Inference: Hybrid SQ-VQ Framework with Adaptive Error Compensation
ICLR 2026Rejected
通讯6
QuantGen: Parameter Generation for Controllable Model Quantization
ICLR 2026Rejected
通讯27
RSAVQ: Riemannian Sensitivity-Aware Vector Quantization for Large Language Models
NeurIPS 2025Poster
通讯21
MambaQuant: Quantizing the Mamba Family with Variance Aligned Rotation Methods
ICLR 2025Poster
23
ASVD: Activation-aware Singular Value Decomposition for Compressing Large Language Models
ICLR 2025Rejected
33
OSTQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
ICLR 2025Poster
三作15
MoEQuant: Enhancing Quantization for Mixture-of-Experts Large Language Models via Expert-Balanced Sampling and Affinity Guidance
ICML 2025Poster
三作26
MQuant: Unleashing the Inference Potential of Multimodal Large Language Models via Full Static Quantization
ICLR 2025Rejected
三作11
RWKVQuant: Quantizing the RWKV Family with Proxy Guided Hybrid of Scalar and Vector Quantization
ICML 2025Poster
通讯29
I-LLM: Efficient Integer-Only Inference for Fully-Quantized Low-Bit Large Language Models
ICLR 2025Rejected
三作