TPS halved after long use with llama.cpp setup on 3070; user reports sudden drop in throughput with same command and models
Read the original at old.reddit.com→previously on my 3070, 32gb ddr4 and i711700 I used this command for months and got 26-30 tps: "C:\Program Files\llama cpp\llama-server.exe" ^ -m "C:\Program Files\llama cpp\models\Qwen3.6-35B-A3B-UD-Q4_K_XL.gguf" ^...
Original headline: "my tps is suddenly halved and I do not know why."
Coverage timeline
- Aug 2, 15:33 UTC r/LocalLLaMA lead source my tps is suddenly halved and I do not know why.