Matrix-CODI on ProsQA shows rank-indifference in latent matrices and a failure of low-rank truncation to hurt accuracy
Read the original at arxiv.org→arXiv:2609.03090v1 Announce Type: new Abstract: Continuous chain-of-thought models compress reasoning into latent tokens. Matrix-valued variants, which route each latent token through a d x d matrix bottleneck,...
Original headline: "The Gradient Does Not See Rank: Rank-Indifference in Matrix-CODI on ProsQA"
Coverage timeline
- Sep 4, 04:00 UTC arXiv cs.LG lead source The Gradient Does Not See Rank: Rank-Indifference in Matrix-CODI on ProsQA