HomePeopleCompaniesAI ModelsOpen SourceAgentsResearchApps
AllModelsInferenceToolingImage and Video
Open SourceTooling 5 Sep 2026 reddit

Benchmarking 21 Qwen3.8 27B quantized variants on 16GB VRAM

The brief

A LocalLLaMA user benchmarked 21 quantized variants of Qwen3.8 27B that fit within 16GB VRAM on an RTX 5080, using real C code as the test task.

Key points

  1. Some quantization levels underperformed expectations, giving practical guidance for users choosing between quants on consumer GPU hardware.
Read the original

Sources