AINewsnow

I thought GPT-4o got 80% faster, but it was just prompt caching messing with my benchmark

This story is from 2026-10-11. It is preserved in the archive; the latest stories are on the live feed.

I got fooled by a benchmark that looked incredible. Same agent workflow. Same model. Same code path. Second run was dramatically faster. My first reaction was the same dumb little hit of engineer dopamine most of us get: nice, we optimized something. We didn’t. The prompt prefix matched, the cache…

Read the full story at DEV Community — AI ↗

Timeline · 1 report

  1. 2026-10-11 06:11 · DEV Community — AI
    I thought GPT-4o got 80% faster, but it was just prompt caching messing with my benchmark

More stories

  1. Introducing GPT-6 in ChatGPT with Intelligent UI — OpenAI YouTube
  2. Which AI (ChatGPT,Claude, Gemini,etc) Is The Best All Around For The Money? — r/ArtificialInteligence
  3. Qwen 3.8 Flash Next is so much fun for three.js — r/LocalLLaMA
  4. A Florida local newspaper ran an op-ed. It was an Iranian AI fake. — Washington Post AI
  5. Anyone else having issues with the ChatGPT app? — r/ChatGPT
  6. OpenAI caught Russians and Iranians using ChatGPT for influence campaigns — NPR Technology
  7. Asana cuts model costs 76x in browser tests with GPT-6.1 Sol — OpenAI News
  8. How Oracle Uses ChatGPT Work to Transform Recruitment — OpenAI YouTube

Get the daily brief of stories like this at 6:30 every morning →