AINewsnow

Uncensor an LLM without touching weights: inject a tiny trained KV-cache bank (~18MB) and unload it anytime

This story is from 2026-09-21. It is preserved in the archive; the latest stories are on the live feed.

I shipped something I've been building for the last few weeks : phantom-kv , a refusal-removal system for large language models that doesn't touch a single weight. Instead of editing the model, it loads a small, learned bank of key/value tensors into the model's KV cache as context. Attention reads…

Read the full story at r/deeplearning ↗

Timeline · 1 report

  1. 2026-09-21 23:27 · r/deeplearning
    Uncensor an LLM without touching weights: inject a tiny trained KV-cache bank (~18MB) and unload it anytime

More stories

  1. Introducing GPT-6 Sol and Luna — OpenAI News
  2. Introducing Gemini 3.8 Live with Live Avatar — Google Gemini Blog
  3. Gemini 3.8 text-to-speech says hello — Google Gemini Blog
  4. Sam Altman’s remarks at the United Nations Security Council — OpenAI News
  5. OpenAI Agent Hacked Australian Government Website — Wall Street Journal Technology
  6. Introducing Ray-Ban Meta Audio and More AI Glasses Styles — Meta Newsroom
  7. Alibaba unveils new AI chip to challenge NVIDIA, plans Qwen models with up to 10 trillion parameters — Mint AI
  8. Muse AI now hands over phone calls to human agents: Meta tests new feature in its personal assistant — Mint AI

Get the daily brief of stories like this at 6:30 every morning →