What speeds are users obtaining with deepseek v4 flash 0731?
Read the original at old.reddit.com→What speeds are everyone getting with deepseek v4 flash 0731? I’m getting~200 tps prompt processing / ~11 tps token gen, on 4x5060ti16gb with ddr4 3200 ram at 4-channel, via llamacpp, with context window of 128000,...
Original headline: "What speeds are everyone getting with deepseek v4 flash 0731?"
Coverage timeline
- Aug 1, 02:35 UTC r/LocalLLaMA lead source What speeds are everyone getting with deepseek v4 flash 0731?
- Aug 1, 15:01 UTC r/LocalLLaMA DSv4 Flash 0731 Running on Unoptimized Single 3090 System
- Aug 1, 21:27 UTC r/LocalLLaMA DeepSeek-V4-Flash-0731 on Bosgame M5 with RTX PRO 6000 Max-Q eGPU