AINewsnow

Two open-weights releases: Victoria (Qwen3.8-Flash-Next with 44% of experts cut, 70% Terminal-Bench 2.1, GGUF included) and Maple (a Canada-first fine-tune)

We had a Dell B300 in the lab for a few weeks and used it to create two fine tunes of Qwen Flash Next. Victoria (coding and agents) Qwen3.8-Flash-Next cut down by 44% using a paper / technique called REAP: 512 down to 288 per layer. Retrained at 4-bit (NVFP4) afterwards, so it's trained for the for…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-09-29 20:33 · r/LocalLLM
    Two open-weights releases: Victoria (Qwen3.8-Flash-Next with 44% of experts cut, 70% Terminal-Bench 2.1, GGUF included) and Maple (a Canada-first fine-tune)

More stories

  1. Qwen 3.8 27B vs Qwen 3.8 Flash Next and time to complete a coding task. — r/LocalLLaMA
  2. Help me plan a Qwen 3.8 Flash Next install on a 5090 + 64gb DDR5 system — r/LocalLLM
  3. Layer Extract & Layer Remove Loras For Qwen Image 2.1 — r/StableDiffusion
  4. I built Slopus, a free, open-source desktop app for generating and editing AI videos locally (Minimax H3) — r/StableDiffusion
  5. Deepseek V4 Flash 0731 on m5 max 128gb — r/LocalLLM
  6. Community reports say the first samples of Qwen 4 are already approaching Fable / Opus-level quality. — r/singularity
  7. Qwen-Image 2.1 Inpainting with LanPaint — alpha channel included — r/StableDiffusion
  8. A LoRA I made: AnyAngle LoRA for Qwen Image 2.1. Style-Aligned Arbitrary Camera Angles — r/StableDiffusion

Get the daily brief of stories like this at 6:30 every morning →