AINewsnow

Qwen 3.8 27B quantized for 16 VRAM or other models (for coding)?

This story is from 2026-09-05. It is preserved in the archive; the latest stories are on the live feed.

Title. I have a 5080 laptop and wondering what is the best way to code locally with 16 VRAM Gemini told me 2 bit quant is not worth it and I'm better off with gpt-oss 20b at 4 bit quant What's your experience?

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-09-05 16:50 · r/LocalLLM
    Qwen 3.8 27B quantized for 16 VRAM or other models (for coding)?

More stories

  1. Pay $39.99 once to put ChatGPT, Claude, Gemini, and more in a single workspace for life — Mashable AI
  2. I gave 6 different AIs the same 5 questions — r/AI_Agents
  3. Alibaba Qwen Releases Qwen3.8-Omni-Flash: A 1M-Context Omni-Modal Model Built Around Agentic Audio-Video Understanding and Tool Use — r/machinelearningnews
  4. My coding agent hit a cold-start 503, found a Gemini key in my repo, and burned $40 while I slept — r/AI_Agents
  5. Gemini 4 Pro vs Fable 5 vs GPT6 Astra — r/GeminiAI
  6. Gemini vs ChatGPT for turning research notes into a presentation structure — r/GeminiAI
  7. ChatGPT vs Gemini — r/GeminiAI
  8. Gemini self-censors in a harmful, obscure way — r/GeminiAI

Get the daily brief of stories like this at 6:30 every morning →