Ling-3.0-tiny 8B A1.3B MoE released with 1.3B active parameters; Ling team highlights performance between 4B and 8-12B Qwen and Gemma models
Read the original at old.reddit.com→Looks like the Ling team open weighted a much smaller version of the Ling-3.0-flash they open weighted a few days ago. It's 8B params with 1.3B active, and seems to fall between the 4B and 8-12B Qwen and Gemma models...
Original headline: "inclusionAI/Ling-3.0-tiny · 8B A1.3B MoE· Hugging Face"
Coverage timeline
- Aug 10, 17:11 UTC r/LocalLLaMA lead source inclusionAI/Ling-3.0-tiny · 8B A1.3B MoE· Hugging Face