qwen3.8-27b-gsq-rco scored very high and fits in 16 GB of VRAM with decent context window — 31.7 tok/s — llm-bench.io
This story is from 2026-09-17. It is preserved in the archive; the latest stories are on the live feed.
We've collected more than 1200 community benchmark results and qwen3.8-27b-gsq-rco just hit one of the highest quality scores we've ever recorded: 88.84/100 , running on a single AMD Radeon RX 9070 (only 16 GB VRAM) at 96k context window (of which max. 38k were used in this run). Of course, at ~30…
Read the full story at r/LocalLLM ↗
Timeline · 2 reports
- 2026-09-17 20:16 · r/huggingface
qwen3.8-27b-gsq-rco scored very high and fits in 16 GB of VRAM with decent context window — 31.7 tok/s — llm-bench.io - 2026-09-17 20:11 · r/LocalLLM
qwen3.8-27b-gsq-rco scored very high and fits in 16 GB of VRAM with decent context window — 31.7 tok/s — llm-bench.io