Qwen-Next seems worse to me then 3.8 27b for coding, but I feel like I must be missing something?
This story is from 2026-09-12. It is preserved in the archive; the latest stories are on the live feed.
Hi! I run both models on MTPLX on my m5 max, and since I have 128GB of ram I run the q8 27b. I think MTPLX only lets me run "optimized for speed" which it says is a dynamic q4 with 8 bit attention. Both of them honestly are very speedy! For coding (in pi agent in nodejs) I've just noticed that 27B…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-09-12 00:05 · r/LocalLLaMA
Qwen-Next seems worse to me then 3.8 27b for coding, but I feel like I must be missing something?