RTX 5090: finding a power-efficiency sweet spot with Qwen3.8-27B
This story is from 2026-08-30. It is preserved in the archive; the latest stories are on the live feed.
I've been playing with q27 / Qwen3.8-27B on my RTX 5090 and got Qwen itself to help me find a reasonable compromise between inference speed and power consumption. Nothing scientific or universal here — just some measurements on my card under Linux. I swept the GPU clock, measured decode tok/s and a…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-08-30 17:49 · r/LocalLLM
RTX 5090: finding a power-efficiency sweet spot with Qwen3.8-27B