Fully quantized NVFP4 Qwen3.8-27B with QUASAR QAD
This story is from 2026-08-26. It is preserved in the archive; the latest stories are on the live feed.
We're releasing a fully quantized NVFP4 version of Qwen3.8-27B. The checkpoint was trained using quantization-aware distillation (QAD) with QUASAR, our new QAT algorithm. We used the original BF16 model as the teacher and distilled the quantized model for 2,446 steps. The checkpoint supports vLLM o…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-08-26 01:04 · r/LocalLLaMA
Fully quantized NVFP4 Qwen3.8-27B with QUASAR QAD