Resource-Efficient pruning for transformer via low-rank importance estimation
Read the original at arxiv.org→arXiv:2608.24973v1 Announce Type: new Abstract: With the rapid development of large-scale pre-trained language models based on Transformer architectures, their high computational and memory costs have become a major...
Original headline: "Resource-Efficient Pruning for Transformer via Low-Rank Importance Estimation"
Coverage timeline
- Aug 27, 04:00 UTC arXiv cs.LG lead source Resource-Efficient Pruning for Transformer via Low-Rank Importance Estimation