Mac with 64GB can run Qwen3.8-Flash-Next-oQ4e-mtp with oMLX on M3Max; user reports 99GB model runs and 130k context window
Read the original at www.reddit.com→I'm genuinely shocked! I was able to run Qwen3.8-Flash-Next-oQ4e-mtp on M3Max 64GB with oMLX! Even a couple of weeks ago, I wasn't able to get it to run. I just tried the latest commit for fun, and it worked! It...
Original headline: "OMG! If you have a Mac with 64GB, try Qwen3.8-Flash-Next-oQ4e-mtp with oMLX!"
Coverage timeline
- Oct 11, 00:00 UTC r/LocalLLaMA lead source OMG! If you have a Mac with 64GB, try Qwen3.8-Flash-Next-oQ4e-mtp with oMLX!