I merged two Qwen3.8-Flash-Next fine-tunes into an experimental GGUF — and it works surprisingly well
I built an experimental Qwen3.8-Flash-Next merge for local inference — and it turned out surprisingly well I've been experimenting with Qwen3.8-Flash-Next and wanted to see what would happen if, instead of simply quantizing another fine-tune, I combined two derivatives with slightly different chara…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-30 18:07 · r/LocalLLM
I merged two Qwen3.8-Flash-Next fine-tunes into an experimental GGUF — and it works surprisingly well