Qwen3.8-Flash-Next (125B) at ~100 tok/s on an M5 Ultra Mac Studio with llama.cpp
Coverage of "Qwen3.8-Flash-Next (125B) at ~100 tok/s on an M5 Ultra Mac Studio with llama.cpp" from 1 source, with a live timeline of who reported what and when.
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-10-10 03:21 · r/LocalLLM
Qwen3.8-Flash-Next (125B) at ~100 tok/s on an M5 Ultra Mac Studio with llama.cpp