I ran Qwen 3.8 27B on my new MacBook Pro M5 Max and on my RTX 5090 workstation. The Mac held up way better than I expected.
This story is from 2026-09-05. It is preserved in the archive; the latest stories are on the live feed.
I have been benchmarking Qwen 3.8 27B (Q4_K_M, llama.cpp) on my RTX 5090 workstation for a while. Last week I ran the exact same sweep on my new MacBook Pro M5 Max, 36GB, just to see how bad the gap would be. The Mac is slower in raw tokens per second. No surprise there, it is a laptop sharing memo…
Read the full story at r/LocalLLM ↗