Codec-Gauge learns equalization transforms for transformer KV-cache compression to improve fidelity
Read the original at arxiv.org→arXiv:2607.20538v1 Announce Type: new Abstract: Long-context Transformer inference increasingly relies on KV-cache compression or quantization. Prior rotation and transform-coding results suggest that the channel...
Original headline: "Codec-Gauge: Learning Compression-Friendly Gauges for Transformer KV Caches"
Coverage timeline
- Jul 24, 04:00 UTC arXiv cs.LG lead source Codec-Gauge: Learning Compression-Friendly Gauges for Transformer KV Caches