Intern S2 Mobius uses Qwen3.5-35B derived model with higher throughput and lower token consumption
Read the original at old.reddit.com→A Qwen3.5-35B derived model with an interesting architectural difference that results in larger throughput and less token consumption (allegedly): https://huggingface.co/internlm/Intern-S2-Mobius submitted by ...
Original headline: "Intern S2 Mobius"
Coverage timeline
- Aug 5, 02:49 UTC r/LocalLLaMA lead source Intern S2 Mobius