GLM 5.3 API Cost: Always Thinking, 2.5x Cheaper per Answer Than 5.2
This story is from 2026-09-15. It is preserved in the archive; the latest stories are on the live feed.
GLM 5.3 costs the same $1.40 per million input tokens and $4.40 per million output as GLM 5.2 , and it no longer lets you turn thinking off: every way of disabling it returns a 400 error, and the default effort is max . On 11 tasks with a checkable answer, run three times each, max was the only set…
Read the full story at DEV Community — AI ↗
Timeline · 1 report
- 2026-09-15 15:22 · DEV Community — AI
GLM 5.3 API Cost: Always Thinking, 2.5x Cheaper per Answer Than 5.2