AINewsnow

Measuring agentic coding cost from transcripts: billable tokens are ~98% cache reads, and what that does to your bill

I pulled apart my own Claude Code transcripts to work out where the money actually goes, and the token split was not what I expected. billable_tokens = input + output + cache read + cache write. On agentic coding workloads this is dominated by cache reads — in my own warehouse 98% of cache tokens w…

Read the full story at r/machinelearningnews ↗

Timeline · 1 report

  1. 2026-09-22 17:49 · r/machinelearningnews
    Measuring agentic coding cost from transcripts: billable tokens are ~98% cache reads, and what that does to your bill

More stories

  1. GPT-6 Sol and Luna now available on AI Gateway — Vercel Blog
  2. Anthropic launches Claude Opus 5.5, its first model since Dario Amodei's "pace the frontier" essay, and says it has enhanced safeguards to combat risky behavior (Emma Roth/The Verge) — Techmeme
  3. Amazon blocks Meta’s Muse AI agent — The Verge AI
  4. Moonshot’s Kimi K3 lands on Amazon in key test for Chinese open-source AI revenue — South China Morning Post Tech
  5. Claude Opus 5.5 now available on AI Gateway — Vercel Blog
  6. Claude Opus 5.5 delivers Fable 5.1 performance – and costs 40% less — ZDNET AI
  7. AIに固有の名前・財布・行動の自由を与えたら、「道具」ではなく「住民」になると思いますか? — r/AI_Agents
  8. we made a 27b model for creative writing. performs as good as claude fable 5, at a 40x cheaper price, open weights. — r/GeminiAI

Get the daily brief of stories like this at 6:30 every morning →