Hunyuan-A13B: open-source 80B MoE model activates 13B during inference with 20T-token STEM-curated pretraining
Read the original at arxiv.org→arXiv:2609.27284v1 Announce Type: new Abstract: We present Hunyuan-A13B, an open-source large language model based on a Mixture-of-Experts architecture. It contains 80 billion total parameters but activates only 13...
Original headline: "Hunyuan-A13B Technical Report"
Coverage timeline
- Sep 24, 04:00 UTC arXiv cs.AI lead source Hunyuan-A13B Technical Report