AINewsnow

[Resource] Z-Image CyberRealistic v9 — Q4_K GGUF that actually fits 8GB VRAM (full bundle, with examples and the x_pad_token fix)

I spent the last few weeks quantizing Z-Image Turbo / CyberRealistic v9 down to GGUF so it runs on 8GB cards without swapping the text encoder to CPU. everything's in one repo — diffusion model, Qwen3-4B encoder, VAE. No hopping between HuggingFace pages. patched quantization to fix ksampler error…

Read the full story at r/comfyui ↗

Timeline · 1 report

  1. 2026-10-06 15:44 · r/comfyui
    [Resource] Z-Image CyberRealistic v9 — Q4_K GGUF that actually fits 8GB VRAM (full bundle, with examples and the x_pad_token fix)

More stories

  1. Is all the work that's being put into Qwen3.8 Flash Next going to set us up for a very quick uplift to Qwen4? — r/LocalLLaMA
  2. Introducing EmbeddingGemma 2: A best-in-class open model for natively multimodal embeddings | Google — r/LocalLLaMA
  3. World Models: The Simulation Strikes Back — r/computervision
  4. I fine-tuned SmolVLM-500M into a lightweight Windows OS Agent (<8GB VRAM) Looking for feedback & ideas! [Weights on HuggingFace] — r/huggingface
  5. ~188k warm ~60–67 tok/s: Qwen3.8-Flash-Next NVFP4 with Strata on a single RTX PRO 4500 32GB + 64GB DDR5. — r/huggingface
  6. A quick Minimax H3 news round-up - 4th October 2026 — r/comfyui
  7. AI Just Crossed the Terrifying Line - Now What? - (Huggingface by Kurzgesagt) — r/ArtificialInteligence
  8. Kurtzgesagt just put up a dive into the dangers of AI agents, looking at the Hugging Face attack. Thoughts? — r/antiai

Get the daily brief of stories like this at 6:30 every morning →