The quant_eval Public Corpus: Behavioral Measurement of Quantization Degradation—Precision, Scale, and Substrate Effects
This paper presents the quant_eval public corpus: a paired behavioral evaluation of full-weight and quantized large language models across eight agent-relevant task families, published as eight open datasets with per-case evidence, paired statistical tests, and a verification chain a reader can check rather than trust....