Fastest qwen 3.8 27b for AMD gpu?
This story is from 2026-08-21. It is preserved in the archive; the latest stories are on the live feed.
Hey, just wondering if there are forks or exact gguf versions that give fastest prompt processing and token gen speeds for AMD gpu? Looking to run q8 or q6 Vram 96gb W7900 + w7800 both 48gb With bandwidth mismatch, tensor paralleling amd equivalent not working
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-08-21 05:47 · r/LocalLLaMA
Fastest qwen 3.8 27b for AMD gpu?