Qwen3.8 27b vs qwen 3.8-flash-next
This story is from 2026-09-05. It is preserved in the archive; the latest stories are on the live feed.
Anyone tried both and seen a diff on performance? Or intelligence? Running 27b on GPU Running Qwen flash on RAM+CPU
Read the full story at r/LocalLLM ↗
Timeline · 6 reports
- 2026-09-08 00:27 · r/LocalLLM
An optimized llama.cpp for people wanting to run Qwen 3.8 Flash Next on two Volta v100 32gbs - 2026-09-07 19:16 · r/LocalLLaMA
exllamav3 comfortably beats llama.cpp running CPU-offloaded Qwen-3.8-Flash-Next on my setup! - 2026-09-07 15:25 · r/LocalLLaMA
Are you running Qwen 3.8 27b or Qwen Flash Next? - 2026-09-06 23:00 · r/LocalLLaMA
vllm + p2p driver hack + qwen 3.8 27B vs llamacpp + qwen flash next ? - 2026-09-06 00:17 · r/LocalLLaMA
Qwen 3.8 Flash Next (Max) is impressive just to talk with. - 2026-09-05 18:50 · r/LocalLLM
Qwen3.8 27b vs qwen 3.8-flash-next