AINewsnow

I built a cache-friendly context compacting plugin for OpenCode

https://github.com/lennartschoch/opencode-cache-compact The default context compacting mechanism in OpenCode strips a bunch of tokens from the beginning of the conversation (system prompt, tools etc). This is fine for hosted models, but on a local model this means you'll prefill the entire conversa…

Read the full story at r/LocalLLaMA ↗

Timeline · 1 report

  1. 2026-09-22 16:32 · r/LocalLLaMA
    I built a cache-friendly context compacting plugin for OpenCode

More stories

  1. Introducing GPT-6 Sol and Luna — OpenAI News
  2. Gemini 3.8 text-to-speech says hello — Google Gemini Blog
  3. Introducing Gemini 3.8 Live with Live Avatar — Google Gemini Blog
  4. Sam Altman’s remarks at the United Nations Security Council — OpenAI News
  5. OpenAI Agent Hacked Australian Government Website — Wall Street Journal Technology
  6. Meta Connect 2026: Muse AI Gadget Shows Ambitious Hardware Vision — Bloomberg AI
  7. Introducing Ray-Ban Meta Audio and More AI Glasses Styles — Meta Newsroom
  8. Muse AI now hands over phone calls to human agents: Meta tests new feature in its personal assistant — Mint AI

Get the daily brief of stories like this at 6:30 every morning →