oMLX, Rapid-MLX, Splash, and MTPLX compared on M3 Max: ~110 tok/s with Qwen3.6-35B-A3B-4bit and ~32 tok/s with Qwen3.8-27B-4bit
Read the original at www.reddit.com→Hello. I picked up a new old stock 14" M3 Max MacBook Pro (14 core CPU / 30 core GPU / 36 GB / 1 TB) from my local market yesterday for around $2,498, and spent the night testing which local inference software is...
Original headline: "oMLX vs Rapid-MLX vs Splash vs MTPLX on M3 Max 36 GB: 110 tok/s on Qwen3.6-35B-A3B, ~32 tok/s on Qwen3.8-27B"
Coverage timeline
- Oct 5, 15:11 UTC r/LocalLLaMA lead source oMLX vs Rapid-MLX vs Splash vs MTPLX on M3 Max 36 GB: 110 tok/s on Qwen3.6-35B-A3B, ~32 tok/s on Qwen3.8-27B