Quants of deepseek flash 0731: recommendations for fast public benchmarks to test quantization effects with about 1 million tokens
Read the original at old.reddit.com→I have various quants of this model and am curious how they perform. can anyone recommend which benchmark would be a good test case for quantization effects? Maybe that can be completed with about 1 million tokens? ...
Original headline: "any reasonably fast public benchmarks I should run quants of deepseek flash 0731 on?"
Coverage timeline
- Aug 8, 21:09 UTC r/LocalLLaMA lead source any reasonably fast public benchmarks I should run quants of deepseek flash 0731 on?