Feeling like I'm not getting enough out of my gpu?
So, for reference, my GPU is an RX 7900 XT, with 32 GB of RAM and a 7800X3D. I'm completely new to local LLMs, but I feel like I'm getting stiffed a little. I see people getting insane context windows while still maintaining great TPS. I'm currently using this model: https://huggingface.co/HauhauCS…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-30 22:38 · r/LocalLLM
Feeling like I'm not getting enough out of my gpu?