Qwen 3.8 Next Flash is really really REALLY verbose..
This story is from 2026-09-07. It is preserved in the archive; the latest stories are on the live feed.
Long time user of 3.6 27b, switched over to Next Flash since it's a logical step up even from 3.8 27b. It's soooo verbose, i'm talking 13 minutes of thinking time on single turn coding requests at approximately 150 tokens per second tg and 7000 tokens per second pp. It's honestly kind of painful to…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-09-07 09:31 · r/LocalLLaMA
Qwen 3.8 Next Flash is really really REALLY verbose..