Qwen Flash Next MTP work restarted
If you're using MTP with Qwen Flash Next and llama.cpp, you can switch to: quants: https://huggingface.co/ggml-org/Qwen3.8-Flash-Next-GGUF PR: https://github.com/ggml-org/llama.cpp/pull/29761 Please note that this is still wip
Read the full story at r/LocalLLaMA ↗
Timeline · 2 reports
- 2026-10-03 01:47 · r/LocalLLM
Strata is seriously impressive, running Qwen 3.8 Flash Next on hermes at 512k context. - 2026-10-01 05:27 · r/LocalLLaMA
Qwen Flash Next MTP work restarted