AINewsnow

Model CPU offload on a 48 GB Mac: 18.5 GB footprint and 9.6 GB of swap

This story is from 2026-10-11. It is preserved in the archive; the latest stories are on the live feed.

The second stageload note ran Qwen-Image-2.1 on a 48 GB Mac two ways. The memory guard stopped eager loading in the VAE decode, and staged loading finished. diffusers has a third way built in, enable_model_cpu_offload , which keeps each model on the CPU and moves it to the GPU only while it runs. O…

Read the full story at DEV Community — Machine Learning ↗

Timeline · 2 reports

  1. 2026-10-11 06:47 · DEV Community — Machine Learning
    How to run Qwen-Image-2.1 on a 48 GB Mac without swap
  2. 2026-10-11 05:44 · DEV Community — Machine Learning
    Model CPU offload on a 48 GB Mac: 18.5 GB footprint and 9.6 GB of swap

More stories

  1. Qwen Image 2.1 Turbo Released -- Hugging Face — r/StableDiffusion
  2. Qwen 3.8 Flash Next is so much fun for three.js — r/LocalLLaMA
  3. Release: Qwen-2B-RCOL Dynamic Low-Bit Quantization (IQ1_M, IQ2_M, IQ3_M) — r/LocalLLaMA
  4. Qwen 3.6 35B A3B: 131K context + vision on 6GB VRAM — r/LocalLLaMA
  5. Hunyuan Image 3 vs Qwen Image 2.1 vs Krea 2 Turbo — r/StableDiffusion
  6. Success with Qwen3.8 27B GSQ-RCO-IQ3_S on 16GB VRAM — r/LocalLLM
  7. Alibaba Qwen Releases Qwen-Image-2.1-Turbo, an 8-Step 7B Image Model — MarkTechPost
  8. Qwen 3.8 27B Q5 vs Qwen 3.8 Next Q3_S for document analysis — r/LocalLLaMA

Get the daily brief of stories like this at 6:30 every morning →