Qwen 3.8 27B - weight vs KV quant
This story is from 2026-09-06. It is preserved in the archive; the latest stories are on the live feed.
Hi, I'm trying to fit Qwen 3.8 onto a 5070 Ti and 5060 Ti 16GB connected via the OCuLink. As I'm chasing near 200k contexts for agentic work, after trying various configurations I've narrowed down my picks to 3 feasible configurations: - Q4_K_XL, 200k context, no KV quantization - Q5_K_M, 200k cont…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-06 11:11 · r/LocalLLM
Qwen 3.8 27B - weight vs KV quant