5070 ti + 64 ram Qwen 3.8 27b
hey everyone, I am getting like 33-40 token speed at 65K Context, do you think this is good numbers? I usually use Qwen 3.8 27b Q4, I try to get better results by edit the bat files assisted by Claude, but I can't go above 40 t/s using LLM CCP, LM Studio.
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-23 16:29 · r/LocalLLM
5070 ti + 64 ram Qwen 3.8 27b