Thank you, Mradermacher; user reports Gemma 4 26B running 75k tokens/s on Hermes with 2x 4060 8GB using lmstudio serving to Hermes
Read the original at www.reddit.com→best iq quants in the biz, got me gemma 4 26b to run 75tok/s tg and 1500 pp on 2x 4060 8gb using lmstudio serving to hermes, much work has been done. submitted by /u/Spiritual_Impress_30 [link] [comments]
Original headline: "Thank You, Mradermacher."
Coverage timeline
- Sep 30, 04:45 UTC r/LocalLLaMA lead source Thank You, Mradermacher.