Worth going from Qwen3.8 27B to flash next or maaybe deepseek v4 flash?
I am running qwen3.8 27b on my dual rtx 3090 (fp8 quant, unquantized cache, 129k context) and I think it works decently well with hermes, opencode etc. But! I am tempted by the new models coming out such as qwen3.8 flash next, deepseek v4 flash, glm 5.3 flash. However, there is a big jump in vram a…
Read the full story at r/LocalLLaMA ↗
Timeline · 2 reports
- 2026-09-27 13:44 · r/LocalLLM
Worth going from Qwen3.8 27B to flash next or maaybe deepseek v4 flash? - 2026-09-27 13:44 · r/LocalLLaMA
Worth going from Qwen3.8 27B to flash next or maaybe deepseek v4 flash?