Toward Extremely Low-Bit and Multi-Precision Conformer and Speech Foundation Model Quantization
Model quantization facilitates Automatic Speech Recognition (ASR) deployment on resource-constrained devices, yet existing methods suffer from severe accuracy loss below 4 bits, redundant storage for multiple precisions, and limited applicability across training paradigms. We address these challenges with two methods t...