Independence prior of SAEs fragments visual concepts
Read the original at arxiv.org→arXiv:2610.04112v1 Announce Type: new Abstract: Sparse Autoencoders (SAEs) decompose model activations into sparse combinations of interpretable dictionary atoms. Although SAEs are grounded in the Linear...
Original headline: "The Independence Prior of SAEs Fragments Visual Concepts"
Coverage timeline
- Oct 6, 04:00 UTC arXiv cs.AI lead source The Independence Prior of SAEs Fragments Visual Concepts