Custom Qwen3.8-27B Quant Hits 99% of BF16 Reasoning at 15% the Size
The brief
A community researcher released a task-aware quantization of Qwen 3.8-27B scoring 82.81% on reasoning benchmarks, against 83.59% for BF16 full precision, at roughly 15% the model size.
Key points
- The approach uses task-aware calibration rather than byte-matched compression, outperforming standard Unsloth UD IQ2_S at the same footprint.
Sources
- redditreddit.com