ESTS submits six entries to the WMT26 Model Compression Shared Task using GPT-OSS-20B with routing-based expert pruning to allocate retained capacity across layers and remove low-importance experts
Read the original at arxiv.org→arXiv:2609.12310v1 Announce Type: new Abstract: We describe six submissions under the team name ESTS to the unconstrained WMT26 Model Compression Shared Task for English--Simplified Chinese and English--Egyptian...
Original headline: "ESTS at WMT26: Routing-Informed Expert Pruning for Model Compression"
Coverage timeline
- Sep 14, 04:00 UTC arXiv cs.CL lead source ESTS at WMT26: Routing-Informed Expert Pruning for Model Compression