GLM-5.3-Flash matches top models at a fraction of the cost, and runs without Nvidia
This story is from 2026-08-27. It is preserved in the archive; the latest stories are on the live feed.
Z.ai releases GLM-5.3-Flash, an open-source model with 320 billion parameters that lands just three points behind the larger GLM-5.3 on Artificial Analysis's Intelligence Index, at a seventh of the cost. What's notable is that all of the inference traffic ran on Chinese AI chips instead of Nvidia h…
Read the full story at The Decoder ↗
Timeline · 1 report
- 2026-08-27 10:24 · The Decoder
GLM-5.3-Flash matches top models at a fraction of the cost, and runs without Nvidia