AINewsnow

Qwen 3.8 27B on a 16GB 5060 Ti and 64gb DDR4. Which quant lands me 10+ tok/s without trashing quality?

This story is from 2026-08-21. It is preserved in the archive; the latest stories are on the live feed.

Trying to settle on the right Qwen 3.8 27B quant for my rig and figured I'd ask people who are actually running it instead of guessing. My setup: GPU: RTX 5060 Ti 16GB (Blackwell) CPU: Ryzen 5 5500 (6c/12t, Zen 3) RAM: 64GB DDR4, dual channel Mobo: B450 micro ATX (so DDR4 plus PCIe 3.0 plus AM4, no…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-08-21 13:24 · r/LocalLLM
    Qwen 3.8 27B on a 16GB 5060 Ti and 64gb DDR4. Which quant lands me 10+ tok/s without trashing quality?

More stories

  1. Alibaba ships Qwen3.8-Omni-Flash to watch, listen and call tools — r/LocalLLM
  2. Alibaba releases Qwen-Image-2.1, a 7B open-weight model it says outperforms most closed-source models, with native transparency and up to ten reference images (Qwen) — Techmeme
  3. Qwen Image 2.1 PR to ComfyUI — r/StableDiffusion
  4. Qwen q4 3.8 27b 16 tok/s 32k RTX 3060 :D — r/LocalLLM
  5. 10 hours left fo Qwen Image 2.1 Public Open Source Release — r/StableDiffusion
  6. Qwen 3.8 27B running on a single RTX 5090 researches and creates a full animation using only code. — r/artificial
  7. US government website used Chinese model the FBI called "malicious" — Ars Technica AI
  8. M2 Mac ultra128gb Qwen flash next — r/LocalLLM

Get the daily brief of stories like this at 6:30 every morning →