AINewsnow

PSA: Use your iGPU for display to save VRAM

I was struggling with a tiny 65k context window on my 4090 running Qwen 3.8 27b with vision & MTP until I remembered I had an iGPU I could be using for Windows. Unplug HDMI from GPU -> plug into motherboard -> free up ~2.5GB of that sweet sweet VRAM! Now I'm on 132k at 125tok/s, with a bit of wiggl…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-10-07 06:47 · r/LocalLLM
    PSA: Use your iGPU for display to save VRAM

More stories

  1. I made a free all-in-one LoRA trainer for consumer GPUs (Windows + Linux): Qwen-Image 2.1, FLUX.2 Klein 9B, Krea 2, Z-Image, Ideogram 4, Anima, SDXL/Pony/Illustrious, LTX 2.3 and MiniMax-H3 (video + audio) — from 4-8 GB VRAM — r/StableDiffusion
  2. Qwen Flash Next on Single B200 or B300, any pointers ? — r/LocalLLM
  3. I mapped every major Qwen release from 2023 to 2026: 44 models, from Qwen-7B to the 2.4T open weights (with sources) — r/machinelearningnews
  4. Qwen3.8-Flash-Next-Q8_0 running on a V100 @ 130Watts 32GB Vram and 128GB System Ram — r/LocalLLM
  5. A benchmark for LLMs playing Civilization V. GLM-5.3 is ahead of Opus-5.5, and Qwen-3.8-27B holds up surprisingly well. — r/LocalLLaMA
  6. My frontier class agent fact-checks my local AI before I grade it. How do you grade your Agents and LLMs? — r/AI_Agents
  7. Update #4: Post training yandex/AliceAI-80B-A3B [instruct!] from scratch — r/LocalLLaMA
  8. Reflection AI Is About to Release a US Open-Weight Model to Take On DeepSeek and Qwen — r/LocalLLaMA

Get the daily brief of stories like this at 6:30 every morning →