Qwen 3.8 2.4T
This story is from 2026-08-22. It is preserved in the archive; the latest stories are on the live feed.
I was able to get Qwen 3.8 2.4T UD-Q1_0 (397gb) to run on quad rtx pro 6000s. Using llama.cpp I was able to fit everything into the GPUs using the following parameters: https://preview.redd.it/nfvish6eaxkh1.png?width=1526&format=png&auto=webp&s=1f6bcbec1713bee98af6f8f05d46b40f21a4dcfe Running nvidi…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-08-22 12:51 · r/LocalLLM
Qwen 3.8 2.4T