Long Responses API runs are costing us more from context than the output itself
I’ve been looking closer at the token usage from one of our longer-running workflows using the Responses API and a big portion of the tokens on the expensive runs aren’t coming from what the model generates. The first couple steps are pretty normal but the workflow uses a few tools and keeps going…
Read the full story at r/OpenAI ↗
Timeline · 1 report
- 2026-10-07 16:22 · r/OpenAI
Long Responses API runs are costing us more from context than the output itself