Qwen3.8 27B Q8_0 @ W7900 - performance question?
Are these reasonable numbers? Qwen3.8 27B Q8_0 Radeon PRO W7900 48GB / ROCm-HIP (gfx1100) FP16 KV / custom tiled GPU attention / all layers GPU DFlash2 Q4_K_M / adaptive drafting, up to 7 tokens / shared head 262,144-token context / 1 concurrent session Razer Core X Chroma external GPU enclosure --…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-10-10 21:37 · r/LocalLLM
Qwen3.8 27B Q8_0 @ W7900 - performance question?