anyone experienced this before?
Before doing anything, my models were running at 20tps. After a few hours has gone by + running and ejecting the model, somehow it magically goes down to 4-5tps. I don't have any program installed, the task manager is fine (no heavy programs running). I was only playing around with deepseek harness…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-25 16:28 · r/LocalLLM
anyone experienced this before?