Test-Time Unlearning via Sparse Autoencoder
This story is from 2026-09-16. It is preserved in the archive; the latest stories are on the live feed.
arXiv:2609.16229v1 Announce Type: new Abstract: Machine unlearning aims to remove specific knowledge from a trained large language model (LLM) without retraining from scratch. Existing methods modify model weights via gradient ascent and its advances. While effective on certain benchmarks, these we…
Read the full story at arXiv cs.LG ↗
Timeline · 1 report
- 2026-09-16 04:00 · arXiv cs.LG
Test-Time Unlearning via Sparse Autoencoder