GLM 5.3 SlopCodeBench Results
This story is from 2026-08-20. It is preserved in the archive; the latest stories are on the live feed.
Howdy once again, I had a request to try out 5.3 on the benchmarks - they're unsaturated so it's a fun test right now! This one was interesting because I accidentally ran it on all 36 problems (rip $200) instead of the 9 i typically do previous runs a b c benchmark context: the ai is tasked to buil…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-08-20 16:04 · r/LocalLLaMA
GLM 5.3 SlopCodeBench Results