AINewsnow

LoRA over GGUF: Train Qwen3.8-Flash-Next in 40G VRAM

https://github.com/woct0rdho/transformers5-qwen3.5-recipe An update to my LoRA over GGUF series: Now we can train Qwen3.8-Flash-Next (125B-A6B + 51B engram) in 40 GiB VRAM, with no CPU offloading, with engram on disk that does not reduce training speed. On Strix Halo it trains context chunk size 20…

Read the full story at r/LocalLLaMA ↗

Timeline · 1 report

  1. 2026-10-09 03:54 · r/LocalLLaMA
    LoRA over GGUF: Train Qwen3.8-Flash-Next in 40G VRAM

More stories

  1. GPT-6 and Intelligent UI for everyone — OpenAI News
  2. Introducing Mistral Large 4 — Mistral AI News
  3. Introducing Claude Haiku 5.5 on AWS — AWS Machine Learning Blog
  4. Sharing AI progress in mathematics — OpenAI News
  5. OpenAI Decisions API now available on AI Gateway — Vercel Blog
  6. Anthropic bans ‘abusive or cruel behavior’ toward Claude — The Verge AI
  7. Introducing Playground: Create and play custom games — Google AI Blog
  8. Anthropic launches OSS Scanner, which provides free, opt-in security audits for open-source projects by sending AI-generated reports without human review (Anthropic) — Techmeme

Get the daily brief of stories like this at 6:30 every morning →