I have a MacBook M4 Pro with 48GB unified memory, anyone running Qwen3.8-27B on a similar config? Looking for some advice on what to run as new to LLM’s (Claude user). It seems that maybe a Q6 quant with MLX and some KV tuning is the way to go for good reasoning?
This story is from 2026-09-09. It is preserved in the archive; the latest stories are on the live feed.
I appreciate it won’t be lightning fast with the GPU bandwidth only being 273 GB/s, but hoping for something usable to reduce my Claude usage? I use it for website design and basic programming and would priortise accuracy over speed as it’s not my day job. I also have a desktop PC with a 5070Ti, am…
Read the full story at r/LocalLLM ↗