HomePeopleCompaniesAI ModelsOpen SourceAgentsResearchApps
AllModelsInferenceToolingImage and Video
Open SourceTooling 8 Sep 2026 reddit

Custom Qwen3.8-27B Quant Hits 99% of BF16 Reasoning at 15% the Size

The brief

A community researcher released a task-aware quantization of Qwen 3.8-27B scoring 82.81% on reasoning benchmarks, against 83.59% for BF16 full precision, at roughly 15% the model size.

Key points

  1. The approach uses task-aware calibration rather than byte-matched compression, outperforming standard Unsloth UD IQ2_S at the same footprint.
Read the original

Sources