AINewsnow

Dwarfstar quants

anyone try these out? https://dwarfstar.sh they have very clever quant techniques, bespoke for a handful of models running on their software. i got qwen 3.8 next running on m3 ultra 96GB studio and its fast and seems good so far. with memory headroom for other stuff kinda blown away to be honest. w…

Read the full story at r/LocalLLaMA ↗

Timeline · 1 report

  1. 2026-10-01 15:12 · r/LocalLLaMA
    Dwarfstar quants

More stories

  1. Two open-weights releases: Victoria (Qwen3.8-Flash-Next with 44% of experts cut, 70% Terminal-Bench 2.1, GGUF included) and Maple (a Canada-first fine-tune) — r/LocalLLM
  2. Qwen flash next on 12+16gb vram, and 32gb ram viable? — r/LocalLLM
  3. Sonnet 5.5 orchestrated a local Qwen 3.8 27B! — r/ClaudeAI
  4. Train Edit Loras for Qwen image 2.1 in Fizgig 6.6.0 — r/StableDiffusion
  5. add GLM-5.3-Flash (GLM5-Next) support by timkhronos · Pull Request #27773 · ggml-org/llama.cpp — r/LocalLLaMA
  6. Continuity update: screen replacement, image to 3D, Qwen Image 2.1, and a lot more since 3.0 — r/StableDiffusion
  7. Is anyone else running insanely long unattended loops? — r/AI_Agents
  8. Viggle turbo v0.3 for Qwen image 2.1: less grain and cleaner surfaces — r/StableDiffusion

Get the daily brief of stories like this at 6:30 every morning →