Optimizing dual AMD r9700 setup
All - I am using the below configuration to generate code and am currently generating ~32 tokens/seconds which is really slow compared to what I think my system could theoretically do. Any idea how to increase token generation rate? Software: OS: Ubuntu 26 GPU Driver: Vulkan LLM runner: llama.cpp L…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-10-02 00:27 · r/LocalLLM
Optimizing dual AMD r9700 setup