RTX-5080 + Qwen 3.8 27B
This story is from 2026-08-22. It is preserved in the archive; the latest stories are on the live feed.
I was able to achieve 17t/s with the uncensored model and use hermes agent as harness with a context window of 64k. Its slower than what im used to but man this is a good model.
Read the full story at r/LocalLLM ↗
Timeline · 2 reports
- 2026-08-24 03:20 · r/LocalLLM
Qwen 3.8-27B on RTX 5080 - 2026-08-22 04:53 · r/LocalLLM
RTX-5080 + Qwen 3.8 27B