InclusionAI Ling-3.0-flash weights released on Hugging Face; MIT reports BF16 and FP8 with 127.5B parameters and 5.1B active users, 512 experts with 8 active per token
Read the original at old.reddit.com→Went public in the last few minutes, both repos ungated. Ling-3.0-flash, BF16, 24 shards, ~255GB Ling-3.0-flash-fp8, official FP8, ~128GB 127.5B total, they quote 5.1B active. What jumped out at me in config.json is...
Original headline: "inclusionAI/Ling-3.0-flash weights are up on Hugging Face — MIT, BF16 plus an official FP8"
Coverage timeline
- Aug 4, 15:21 UTC r/LocalLLaMA lead source inclusionAI/Ling-3.0-flash weights are up on Hugging Face — MIT, BF16 plus an official FP8