Model Releases

any reasonably fast public benchmarks I should run quants of deepseek flash 0731 on?

I have various quants of this model and am curious how they perform. can anyone recommend which benchmark would be a good test case for quantization effects? Maybe that can be completed with about 1 m

DGX agentreddit
model-releasesr-localllama

I have various quants of this model and am curious how they perform. can anyone recommend which benchmark would be a good test case for quantization effects? Maybe that can be completed with about 1 million tokens? submitted by /u/nomorebuttsplz [link] [comments]

Related

Source: r/LocalLLaMA | 2026-08-08

Loading related sources…