Test-time unlearning via sparse autoencoder
Read the original at arxiv.org→arXiv:2609.16229v1 Announce Type: new Abstract: Machine unlearning aims to remove specific knowledge from a trained large language model (LLM) without retraining from scratch. Existing methods modify model weights...
Original headline: "Test-Time Unlearning via Sparse Autoencoder"
Coverage timeline
- Sep 16, 04:00 UTC arXiv cs.LG lead source Test-Time Unlearning via Sparse Autoencoder