AINewsnow

I posit QWEN team will dust off the old 397B-A17B architecture to compete with Deepseek V4 0731 Flash

This story is from 2026-08-19. It is preserved in the archive; the latest stories are on the live feed.

To me, it just makes rational business sense. DS4 is king on openrouter and has been pretty much since the day it was launched. It's size, cost and intelligence seems to be the sweet spot for current developer requirements. Qwen does have the old 235B-A22B architecture under their belt as well but…

Read the full story at r/LocalLLaMA ↗

Timeline · 1 report

  1. 2026-08-19 16:03 · r/LocalLLaMA
    I posit QWEN team will dust off the old 397B-A17B architecture to compete with Deepseek V4 0731 Flash

More stories

  1. A company ran 8 identical AI societies for weeks with different models and just published what happened. Some of it is genuinely unsettling. — r/ArtificialInteligence
  2. Alibaba ships Qwen3.8-Omni-Flash to watch, listen and call tools — r/LocalLLM
  3. Cactus Needle 3: A Sliceable 8-29MB Automation Foundation Model That Matches DeepSeek v4 Flash — r/LocalLLaMA
  4. Qwen 3.8 27B Running for 63 hours on a RTX 3090 to solve the Riemann hypothesis — r/LocalLLM
  5. 10 hours left fo Qwen Image 2.1 Public Open Source Release — r/StableDiffusion
  6. US government website used Chinese model the FBI called "malicious" — Ars Technica AI
  7. Qwen q4 3.8 27b 16 tok/s 32k RTX 3060 :D — r/LocalLLM
  8. Deployed Qwen 3.6 35B A3B on a single DGX Spark supporting 12 concurrent users at 262K context. Are there better ways to optimize this? — r/LocalLLM

Get the daily brief of stories like this at 6:30 every morning →