Help Optimizing Q4 Qwen3.8 27b fully on RX 6800xt (beellama.cpp)
This story is from 2026-09-03. It is preserved in the archive; the latest stories are on the live feed.
Hi there, I've done a lot of research seen people successfully running Qwen3.8 27b on 16gb cards, but I'm not seeing similar results and would like some help optimizing. I am currently running empero-ai/Qwen3.8-27B-Ridge-GGUF using beellama.cpp, mmproj offloaded to CPU. Theoretically, everything sh…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-03 21:51 · r/LocalLLM
Help Optimizing Q4 Qwen3.8 27b fully on RX 6800xt (beellama.cpp)