AINewsnow

Strata Qwen 3.8 flash next is the biggest thing since the release of Qwen 3.8 27b

It provides the speed and somewhat accessibility to low vram users to run Flash Next as their main. It's providing the speed of qwen 3.5-35-a3b but with higher intelligence over 3.8 27b. This will allow many people to replace 27b. I'm getting around 80-110 tok/s 2500 pp. 2x 3090's 96gb ddr5 ram. I'…

Read the full story at r/LocalLLaMA ↗

Timeline · 12 reports

  1. 2026-10-06 14:58 · r/LocalLLaMA
    Qwen3.8-Flash-Next on Strata
  2. 2026-10-06 10:37 · r/LocalLLaMA
    NInfer6000 - Qwen 3.8 Flash Next @ 400 tg/s & 13K pp/s
  3. 2026-10-05 17:38 · r/LocalLLM
    Strata: Qwen 3.8 Flash next in loop
  4. 2026-10-05 12:28 · r/LocalLLM
    Strata for Windows/AMD GPUs, Qwen 3.8 Flash Next with large (128K+) context coding performance
  5. 2026-10-05 10:50 · r/LocalLLaMA
    MoE SSD streaming on a 64 GB Mac mini: GPU still waits 27% of decode on experts. Ideas?
  6. 2026-10-05 07:27 · r/LocalLLM
    4x 3080 20GB (modded, alibaba) + Strata (Qwen Flash Next 125B IQ3_XXS) = 105 t/s generation 5000 prompt processing on 150k context
  7. 2026-10-05 07:21 · r/LocalLLM
    Upgrade existing PC to run Qwen 3.8 Flash Next via Strata or swap to a Strix Halo?
  8. 2026-10-05 06:44 · r/LocalLLaMA
    I got Qwen Flash Next Q4 running on a Mac Mini m5 64gb with ssd streaming
  9. 2026-10-05 06:35 · r/LocalLLM
    Got Qwen Flash Next Q4 running on my Mac Mini M5 64GB with ssd streaming
  10. 2026-10-05 06:10 · r/LocalLLM
    Qwen Flash Next on Single B200 or B300, any pointers ?
  11. 2026-10-04 02:48 · r/LocalLLM
    Anyone tried strata qwen.38 flash next on a RX6700XT?
  12. 2026-10-03 23:44 · r/LocalLLaMA
    Strata Qwen 3.8 flash next is the biggest thing since the release of Qwen 3.8 27b

More stories

  1. The Story of Qwen: Alibaba's AI Models From 7B to 2.4T — MarkTechPost
  2. I made a free all-in-one LoRA trainer for consumer GPUs (Windows + Linux): Qwen-Image 2.1, FLUX.2 Klein 9B, Krea 2, Z-Image, Ideogram 4, Anima, SDXL/Pony/Illustrious, LTX 2.3 and MiniMax-H3 (video + audio) — from 4-8 GB VRAM — r/StableDiffusion
  3. Is all the work that's being put into Qwen3.8 Flash Next going to set us up for a very quick uplift to Qwen4? — r/LocalLLaMA
  4. A benchmark for LLMs playing Civilization V. GLM-5.3 is ahead of Opus-5.5, and Qwen-3.8-27B holds up surprisingly well. — r/LocalLLaMA
  5. My frontier class agent fact-checks my local AI before I grade it. How do you grade your Agents and LLMs? — r/AI_Agents
  6. ComfyUI Qwen image 2.1 Enhancer (Two nodes) — r/StableDiffusion
  7. Update #4: Post training yandex/AliceAI-80B-A3B [instruct!] from scratch — r/LocalLLaMA
  8. Reflection AI Is About to Release a US Open-Weight Model to Take On DeepSeek and Qwen — r/LocalLLaMA

Get the daily brief of stories like this at 6:30 every morning →