Qwen3.8 Flash Next - Strix Halo
This story is from 2026-09-07. It is preserved in the archive; the latest stories are on the live feed.
So i have been running qwen3.8 flash next ud q4 k xl at 256k context getting at the start aroudn 270pp and 21tg with mtp with qwen 27b ud q3 k xl 96k context on the 9060xt as a callable subagent and the one theing that i am really liking about this model is it doesnt stop and wait for me to contino…
Read the full story at r/LocalLLM ↗
Timeline · 2 reports
- 2026-09-09 01:34 · r/LocalLLaMA
Qwen3.8-Flash-Next on MLX-serve, 1m context is released! - 2026-09-07 10:43 · r/LocalLLM
Qwen3.8 Flash Next - Strix Halo