Qwen4’s architecture is here early, firing 6B parameters out of 125B
This story is from 2026-08-26. It is preserved in the archive; the latest stories are on the live feed.
Alibaba’s Qwen team has released Qwen3.8-Flash-Next, an open-weight preview of the architecture it intends to use for Qwen4, carrying 125B parameters but activating only 6B for each token. Its licence may not qualify for the EU AI Act’s open-source exemption. Alibaba’s Qwen team has published the a…
Read the full story at The Next Web ↗
Timeline · 1 report
- 2026-08-26 19:55 · The Next Web
Qwen4’s architecture is here early, firing 6B parameters out of 125B