M2 Mac ultra128gb Qwen flash next
I am trying to get good speeds for Qwen flash next but I also need some ram for the system. Right now gpu is at 96gb around more I cannot use. Before I was getting 50 token/s from omlx and 60 tokens/s from mlx serve but lower context like 32k. I had ngram on ssd now where my system has more ram for…
Read the full story at r/LocalLLM ↗
Timeline · 3 reports
- 2026-09-20 18:16 · r/LocalLLM
I ran claude vs Qwen 3.8 Flash Next - 2026-09-19 22:55 · r/LocalLLaMA
Qwen-3.8-Flash-Next on 1x RTX 5090: TG=50 t/s, PP=2300 t/s - with FreeToken - 2026-09-19 01:20 · r/LocalLLM
M2 Mac ultra128gb Qwen flash next