Qwen 3.8 27b shows 22 t/s performance with Q4, max context, llama.cpp, and MTP enabled; user reports and configurations posted on Reddit
Read the original at old.reddit.com→Looking for your experiences, your speeds, and your configs. I myself am getting an abysmal 22t/s with the q4, max context, llama.cpp, MTP enabled. submitted by /u/Acrobatic_Stress1388 [link] [comments]
Original headline: "Users of Qwen 3.8 27b on the strict halo, REPORT!"
Coverage timeline
- Aug 16, 06:55 UTC r/LocalLLaMA lead source Users of Qwen 3.8 27b on the strict halo, REPORT!