What KV cache headroom would I get with 2x AMD R9700 running Qwen 3.8-27b?
I'm not really familiar with LLM/SLM deployment, I was waiting a bit to see whether new architectures (especially Mamba/RWKV) would give significantly more intelligence per compute. I've recently read a lot around the R9700, and I'd love to experiment with it. But buying two R9700 is quite pricey,…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-18 20:19 · r/LocalLLM
What KV cache headroom would I get with 2x AMD R9700 running Qwen 3.8-27b?