DecodeShare identifies a shared low-dimensional subspace in decode-time hidden states and tests its causal role by removing it during decoding
Read the original at arxiv.org→arXiv:2607.20469v1 Announce Type: new Abstract: Large language models (LLMs) handle many tasks with one set of parameters, but under KV-cached inference it is unclear what task-general structure, if any, is used at...
Original headline: "DecodeShare: Tracing the Shared Subspace of LLM Decode-Time Decisions"
Coverage timeline
- Jul 24, 04:00 UTC arXiv cs.AI lead source DecodeShare: Tracing the Shared Subspace of LLM Decode-Time Decisions