Manifold projection and iterative autoencoder refinement for masked language modeling
Read the original at arxiv.org→arXiv:2609.30288v1 Announce Type: new Abstract: In Transformer-based masked language models, attention is the primary mechanism for context mixing, but there are other ways to mix data across tokens. Recent...
Original headline: "Manifold Projection and Iterative Autoencoder Refinement for Masked Language Modeling"
Coverage timeline
- Sep 28, 04:00 UTC arXiv cs.CL lead source Manifold Projection and Iterative Autoencoder Refinement for Masked Language Modeling