Qwen 3.6 27B flags/settings in llama.cpp
Read the original at old.reddit.com→I run the following on a 5090 and have been okay with its performance, it does most things somewhere 80-100 t/s, though that can slow down at full 262k context - more like 40 t/s at times. I use it primarily in...
Coverage timeline
- Aug 7, 21:25 UTC r/LocalLLaMA lead source Qwen 3.6 27B flags/settings in llama.cpp