Questions about qwen 3.8 27b mlx on M5 pro 64G model
This story is from 2026-09-03. It is preserved in the archive; the latest stories are on the live feed.
I heard that token genertion was improved to 20+ token/s , but it was still 17 token/s as same as qwen3.6. What should I do to get that boost? I tested qwen3.8 27b 4bit model on omlx and lm studio, still no difference.
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-03 20:05 · r/LocalLLM
Questions about qwen 3.8 27b mlx on M5 pro 64G model