What model sits between Qwen 3.8 27b and Flash next for coding?
Having tested both Qwen 3.8 27b and Flash next on RTX 5090 with 96GB RAM, I want to find the middle ground between the two for coding capabilities but not sacrifice decode speed to standstill. I would like decode speed to between 75-100 ideally for fast iterations; otherwise I become impatient. Cur…
Read the full story at r/LocalLLaMA ↗
Timeline · 3 reports
- 2026-09-29 11:18 · r/LocalLLaMA
Qwen 3.8 27B Q4 with 100K context on a 16 GB RX 7800 XT guide - 2026-09-28 16:25 · r/LocalLLM
Qwen 3.8 27B Q4/Q6/Q8 vs Qwen 3.8 Flash-Next on a 96GB M2 Max - 2026-09-28 13:25 · r/LocalLLaMA
What model sits between Qwen 3.8 27b and Flash next for coding?