Qwen 3.8 27B quantized for 16 VRAM or other models (for coding)?
This story is from 2026-09-05. It is preserved in the archive; the latest stories are on the live feed.
Title. I have a 5080 laptop and wondering what is the best way to code locally with 16 VRAM Gemini told me 2 bit quant is not worth it and I'm better off with gpt-oss 20b at 4 bit quant What's your experience?
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-05 16:50 · r/LocalLLM
Qwen 3.8 27B quantized for 16 VRAM or other models (for coding)?