Unsloth's IQ3_S quant of Qwen 3.8 27b is insane
This story is from 2026-08-30. It is preserved in the archive; the latest stories are on the live feed.
For context, I have a 16 GB VRAM card (RX 9060 XT) - I am writing a kinda complex project in Rust. Unsloth's IQ3_S quant of Qwen 3.8 27b is the best thing I can fit into VRAM at a reasonable context length. And despite the aggressive quantization, the model works perfectly. It doesn't make ownershi…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-08-30 23:08 · r/LocalLLM
Unsloth's IQ3_S quant of Qwen 3.8 27b is insane