WinterMix releases Qwen3.5-122B-A10B native MLX builds at 82 GiB and 68 GiB for agent swarms, claiming the 82 GiB build outperforms 94–95 GiB 6-bit quantizations while staying native MLX.
Read the original at old.reddit.com→TL;DR: I spent 9 days developing a new quantization method for MLX models and measured 18 variants against each other on a single M5 Max MacBook Pro (128 GB). The result is the best-measuring MLX quant of...
Original headline: "[Release] WinterMix — Qwen3.5-122B-A10B in native MLX: an 82 GiB build that beats 94–95 GiB quants, plus a 68 GiB build for agent swarms"
Coverage timeline
- Aug 2, 08:42 UTC r/LocalLLaMA lead source [Release] WinterMix — Qwen3.5-122B-A10B in native MLX: an 82 GiB build that beats 94–95 GiB quants, plus a 68 GiB build for agent swarms