My RULE of Thumb of choosing a models
This story is from 2026-09-03. It is preserved in the archive; the latest stories are on the live feed.
This is mostly for setting up for expectation, since personally without LLM i could take 3 days (15 hours of active programming) to debug or implement a feature, but with Qwen 27B (even before Qwen 3.8) it take 4 hours. And yes 0.5 tok/s is human, not accounting of deletion and pausing, that's also…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-09-03 06:27 · r/LocalLLaMA
My RULE of Thumb of choosing a models