Glyph: I made two AI models fight inside real terminals and the only honest judge is a PTY
This story is from 2026-09-24. It is preserved in the archive; the latest stories are on the live feed.
Model benchmarks are numbers on a leaderboard. Nobody feels them. I wanted the opposite: send one prompt to two models, watch both answers animate live in terminals next to each other , and keep a replayable artifact as proof. That is Glyph — a "model duel" app where the scoreboard is two CRT panes…
Read the full story at DEV Community — AI ↗
Timeline · 1 report
- 2026-09-24 01:56 · DEV Community — AI
Glyph: I made two AI models fight inside real terminals and the only honest judge is a PTY