Qwen3.8 27b Ninfer Windows Edition x 5090
This story is from 2026-09-17. It is preserved in the archive; the latest stories are on the live feed.
Pretty pleased with the results. I was running LM Studio Q6 before and after this switch to ninfer my decode and prompt speeds effectively doubled. Decode / generation: ~162.1 tok/s Prefill: ~4.26k tok/s 200k Context Q8
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-17 00:57 · r/LocalLLM
Qwen3.8 27b Ninfer Windows Edition x 5090