AINewsnow

Is all the work that's being put into Qwen3.8 Flash Next going to set us up for a very quick uplift to Qwen4?

Given the commentary on the Q3.8FN release page here https://qwen.ai/blog?id=qwen3.8-flash-next I assume/hope that all the work that's going on to optimise the hell out of running it will be useful when Qwen4 drops?

Read the full story at r/LocalLLaMA ↗

Timeline · 2 reports

  1. 2026-10-04 05:40 · r/LocalLLM
    Tesla V100 32GB, RTX3090, Strata Qwen3.8-Flash-Next and Qwen3.8-27B playing around
  2. 2026-10-04 01:52 · r/LocalLLaMA
    Is all the work that's being put into Qwen3.8 Flash Next going to set us up for a very quick uplift to Qwen4?

More stories

  1. Strata is seriously impressive, running Qwen 3.8 Flash Next on hermes at 512k context. — r/LocalLLM
  2. Pi extension: Skip reasoning with local Qwen 27B and proceed to answer right now — r/LocalLLaMA
  3. Qwen4Exp: add MTP by am17an · Pull Request #29761 · ggml-org/llama.cpp — r/LocalLLaMA
  4. I built Ninfer 4080 for 16GB class GPUs — r/LocalLLaMA
  5. Use Qwen-Image-2.1-viggle-turbo to generate character sheets in ComfyUI — r/StableDiffusion
  6. Direct weight surgery from Qwen-4B to 0.8B on an 8GB RX 580: why editing all layers breaks everything, and how 4 anchor blocks fixed it — r/machinelearningnews
  7. I made my iPhone a second GPU for my 24 GB MacBook: Qwen 3.8 27B prefills 29–44% faster & my holds part of the CTX window. — r/LocalLLaMA
  8. What is your experience with bonsai 2 27b? — r/ArtificialInteligence

Get the daily brief of stories like this at 6:30 every morning →