AINewsnow

Success with Qwen3.8 27B GSQ-RCO-IQ3_S on 16GB VRAM

Took me a lot of reading and testing to get Qwen3.8 up and running and producing very good coding/agent results, so maybe it is helpful for others if I share my setup. It fits nicely in VRAM (max utilization is about 15.2GB). Performance starts at about 600t prefill and 26t/s, then degrades to 22t/…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-10-08 22:09 · r/LocalLLM
    Success with Qwen3.8 27B GSQ-RCO-IQ3_S on 16GB VRAM

More stories

  1. Introducing Mistral Large 4 — Mistral AI News
  2. Introducing Claude Haiku 5.5 on AWS — AWS Machine Learning Blog
  3. GPT-6 and Intelligent UI for everyone — OpenAI News
  4. Sharing AI progress in mathematics — OpenAI News
  5. Anthropic launches OSS Scanner, which provides free, opt-in security audits for open-source projects by sending AI-generated reports without human review (Anthropic) — Techmeme
  6. OpenAI Decisions API now available on AI Gateway — Vercel Blog
  7. Anthropic bans ‘abusive or cruel behavior’ toward Claude — The Verge AI
  8. Introducing Playground: Create and play custom games — Google AI Blog

Get the daily brief of stories like this at 6:30 every morning →