AINewsnow

DeepSeek V4 0731 -> Qwen 3.8 Flash -> GLM 5.3 Flash (and back again!)

This story is from 2026-08-27. It is preserved in the archive; the latest stories are on the live feed.

Spent yesterday getting Qwen3.8 Flash and GLM 5.3 Flash up and running on my cluster of 4 x DGX Sparks with a view to replacing DeepSeek 0731... but.. really not that impressed with GLM 5.3 - overly verbose and takes for ever (was getting around 22 tok/s on dual spark setup). Have now got myself se…

Read the full story at r/LocalLLaMA ↗

Timeline · 1 report

  1. 2026-08-27 11:05 · r/LocalLLaMA
    DeepSeek V4 0731 -> Qwen 3.8 Flash -> GLM 5.3 Flash (and back again!)

More stories

  1. I gave 6 different AIs the same 5 questions — r/AI_Agents
  2. Alibaba Qwen Releases Qwen3.8-Omni-Flash: A 1M-Context Omni-Modal Model Built Around Agentic Audio-Video Understanding and Tool Use — r/machinelearningnews
  3. Qwen Image 2.1 PR to ComfyUI — r/StableDiffusion
  4. Alibaba releases Qwen-Image-2.1, a 7B open-weight model it says outperforms most closed-source models, with native transparency and up to ten reference images (Qwen) — Techmeme
  5. Bolt Adds DeepSeek V4.1 Flash at 10x Cheaper Than V4 Pro — AlphaSignal
  6. Qwen q4 3.8 27b 16 tok/s 32k RTX 3060 :D — r/LocalLLM
  7. 10 hours left fo Qwen Image 2.1 Public Open Source Release — r/StableDiffusion
  8. Qwen 3.8 27B running on a single RTX 5090 researches and creates a full animation using only code. — r/artificial

Get the daily brief of stories like this at 6:30 every morning →