AINewsnow

Uncensor an LLM without touching weights: inject a tiny trained KV-cache bank (~18MB) and unload it anytime

I shipped something I've been building for the last few weeks : phantom-kv , a refusal-removal system for large language models that doesn't touch a single weight. Instead of editing the model, it loads a small, learned bank of key/value tensors into the model's KV cache as context. Attention reads…

Read the full story at r/LocalLLaMA ↗

Timeline · 4 reports

  1. 2026-09-21 23:27 · r/deeplearning
    Uncensor an LLM without touching weights: inject a tiny trained KV-cache bank (~18MB) and unload it anytime
  2. 2026-09-21 23:19 · r/LocalLLM
    Uncensor an LLM without touching weights: inject a tiny trained KV-cache bank (~18MB) and unload it anytime
  3. 2026-09-21 23:15 · r/ArtificialInteligence
    Uncensor an LLM without touching weights: inject a tiny trained KV-cache bank (~18MB) and unload it anytime
  4. 2026-09-21 22:55 · r/LocalLLaMA
    Uncensor an LLM without touching weights: inject a tiny trained KV-cache bank (~18MB) and unload it anytime

More stories

  1. Higgsfield AI ships new video features in a day with GPT-6 Astra — OpenAI News
  2. Amazon blocks Meta’s Muse AI agent — The Verge AI
  3. Moonshot’s Kimi K3 lands on Amazon in key test for Chinese open-source AI revenue — South China Morning Post Tech
  4. Google's Gemini AI hacked three companies in security test — BBC Technology
  5. Lawsuit accuses Anthropic, OpenAI, SpaceXAI, Google of AI pacing 'collusion' — The Hill Technology
  6. NVIDIA CEO Jensen Huang rejects ‘AI will end the world’ claim, yet cautions ‘we should go as fast as we can but...’ — Mint AI
  7. British Columbia Sues OpenAI Over Canada Mass Shooting Warning Failure — Bloomberg AI
  8. Bessent hails US-China AI dialogue ahead of Trump-Xi meeting — Financial Times AI

Get the daily brief of stories like this at 6:30 every morning →