I thought GPT-4o got 80% faster, but it was just prompt caching messing with my benchmark
This story is from 2026-10-11. It is preserved in the archive; the latest stories are on the live feed.
I got fooled by a benchmark that looked incredible. Same agent workflow. Same model. Same code path. Second run was dramatically faster. My first reaction was the same dumb little hit of engineer dopamine most of us get: nice, we optimized something. We didn’t. The prompt prefix matched, the cache…
Read the full story at DEV Community — AI ↗
Timeline · 1 report
- 2026-10-11 06:11 · DEV Community — AI
I thought GPT-4o got 80% faster, but it was just prompt caching messing with my benchmark