Qwen3.6-35B-A3B on RX 7800 XT (16GB VRAM) — 33 t/s at long context
This story is from 2026-09-02. It is preserved in the archive; the latest stories are on the live feed.
I've been running Qwen3.6-35B-A3B locally on an AMD RX 7800 XT (16GB VRAM, 32GB RAM) and wanted to share the full setup since I spent weeks fighting the same problems everyone else is. TL;DR: The model works great, but the default config is a trap. If you just offload layers to GPU and keep the KV…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-02 09:24 · r/LocalLLM
Qwen3.6-35B-A3B on RX 7800 XT (16GB VRAM) — 33 t/s at long context