SPIN: Shadow Predictive Indexer reduces indexer overhead in sparse attention via history-based KV block prediction
Read the original at arxiv.org→arXiv:2610.09025v1 Announce Type: new Abstract: Indexer-based sparse attention reduces the cost of core attention by passing only a fixed, small number of important tokens to it. However, the indexer must still...
Original headline: "SPIN: Shadow Predictive Indexer for Sparse Attention"
Coverage timeline
- Oct 8, 04:00 UTC arXiv cs.LG lead source SPIN: Shadow Predictive Indexer for Sparse Attention