AINewsnow

Qwen4’s architecture is here early, firing 6B parameters out of 125B

This story is from 2026-08-26. It is preserved in the archive; the latest stories are on the live feed.

Alibaba’s Qwen team has released Qwen3.8-Flash-Next, an open-weight preview of the architecture it intends to use for Qwen4, carrying 125B parameters but activating only 6B for each token. Its licence may not qualify for the EU AI Act’s open-source exemption. Alibaba’s Qwen team has published the a…

Read the full story at The Next Web ↗

Timeline · 1 report

  1. 2026-08-26 19:55 · The Next Web
    Qwen4’s architecture is here early, firing 6B parameters out of 125B

More stories

  1. Alibaba ships Qwen3.8-Omni-Flash to watch, listen and call tools — r/LocalLLM
  2. Alibaba releases Qwen-Image-2.1, a 7B open-weight model it says outperforms most closed-source models, with native transparency and up to ten reference images (Qwen) — Techmeme
  3. Qwen Image 2.1 PR to ComfyUI — r/StableDiffusion
  4. Qwen q4 3.8 27b 16 tok/s 32k RTX 3060 :D — r/LocalLLM
  5. 10 hours left fo Qwen Image 2.1 Public Open Source Release — r/StableDiffusion
  6. Qwen 3.8 27B running on a single RTX 5090 researches and creates a full animation using only code. — r/artificial
  7. US government website used Chinese model the FBI called "malicious" — Ars Technica AI
  8. M2 Mac ultra128gb Qwen flash next — r/LocalLLM

Get the daily brief of stories like this at 6:30 every morning →