Qwen 35B-A3B MoE is ~4× faster than Qwen 27B dense on local coding tests with a smaller quality gap than expected
Read the original at old.reddit.com→I compared Qwen 35B-A3B MoE against Qwen 27B dense on a series of local coding-maintenance tasks. On my R9700/llama.cpp setup, the MoE model generated about 3.9× faster (~116 vs ~30 tok/s), but the coding-quality...
Original headline: "Qwen 35B-A3B MoE vs 27B dense in local coding tests: ~4× faster, much smaller quality gap than I expected"
Coverage timeline
- Aug 8, 05:44 UTC r/LocalLLaMA lead source Qwen 35B-A3B MoE vs 27B dense in local coding tests: ~4× faster, much smaller quality gap than I expected