Run Qwen3.8+Flash-Next and tiny models on Apple Silicon up to 3x faster
Maybe you'll like it? I hope I get to use my self-promotion credit a tiny little bit here after being in the community so long haha. I was the top of MLX.fast for a while and remain the winner on chips below M5. If you have capacity to contribute further enhancements I'd love that https://github.co…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-09-26 21:10 · r/LocalLLaMA
Run Qwen3.8+Flash-Next and tiny models on Apple Silicon up to 3x faster