MoE$^2$-LoRA: when MoE models meet MoE-style low-rank adaptation
Read the original at arxiv.org→arXiv:2607.21978v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures have been widely adopted in large language models, yet parameter-efficient fine-tuning (PEFT) for MoE models remains...
Original headline: "MoE$^2$-LoRA: When MoE Models Meet MoE-style Low-Rank Adaptation"
Coverage timeline
- Jul 27, 04:00 UTC arXiv cs.CL lead source MoE$^2$-LoRA: When MoE Models Meet MoE-style Low-Rank Adaptation