AINewsnow

Surgical Subgraphs: How We Cut Coding-Agent Token Costs by 95%

This story is from 2026-09-29. It is preserved in the archive; the latest stories are on the live feed.

Every coding agent has the same expensive habit: when it needs context, it grabs too much of it. A full-file dump here, a 50-chunk RAG retrieval there — and suddenly a simple "trace this API call" task burns 12,000+ tokens before the model writes a single line of code. At Claude 3.5 Sonnet pricing…

Read the full story at DEV Community — AI ↗

Timeline · 1 report

  1. 2026-09-29 09:45 · DEV Community — AI
    Surgical Subgraphs: How We Cut Coding-Agent Token Costs by 95%

More stories

  1. Introducing Claude Sonnet 5.5 on AWS — AWS Machine Learning Blog
  2. Anthropic warns of ‘existential risks to humanity’ in IPO prospectus — Financial Times AI
  3. Opus 5.5 — r/ClaudeAI
  4. Sonnet 5.5 — Hacker News Front Page
  5. Kimi K3: A Claude clone or something else? — CoreWeave Blog
  6. Tutorial: Benchmarking GPT-6 Astra vs Claude Fable 5.1 vs GPT-5.6 Sol using W&B Weave — CoreWeave Blog
  7. Claude Sonnet 5.5 now available on AI Gateway — Vercel Blog
  8. Anthropic releases Claude Sonnet 5.5, a faster and cheaper follow-up to Opus 5.5 — Mashable AI

Get the daily brief of stories like this at 6:30 every morning →