Are you running Qwen 3.8 27b or Qwen Flash Next?
This story is from 2026-09-07. It is preserved in the archive; the latest stories are on the live feed.
Curious about what people are preferring, if you have the hardware. I have m3 Max 96gb and both run, and largely feel identical, but prefill on qwen 27b is faster. Is there anything / anyone working on anything to improve pp with mlx? Branching question: is anyone working on a harness that works wi…
Read the full story at r/LocalLLaMA ↗
Timeline · 3 reports
- 2026-09-09 23:42 · r/LocalLLaMA
Running qwen 3.8 27B iq3 xxs on RTX 3060. - 2026-09-08 15:53 · r/LocalLLM
Running Qwen 3.8 27B at Q4 on 16GB VRAM at 200K CTX at 50t/s - 2026-09-07 15:25 · r/LocalLLaMA
Are you running Qwen 3.8 27b or Qwen Flash Next?