Model Releases
Ollama Cloud Quota Benchmark
Recently I bought an Ollama Cloud sub and accidently spent my whole 5h quota upon using DeepSeek V4 Pro... but why? isnt it supposed to be a cheap model? Youd think there would be a correlation betwee
Recently I bought an Ollama Cloud sub and accidently spent my whole 5h quota upon using DeepSeek V4 Pro... but why? isnt it supposed to be a cheap model? Youd think there would be a correlation between quota usage and intelligence of the model being used. After doing my own benchmark with all models, it becomes pretty clear that there is not. If you want to read more about it: https://blog.micr.dev/blog/i-made-my-first-benchmark Or if you want to look into the eval and run it yourself: https://github.com/Microck/ollama-quota-bench submitted by /u/MicrockYT [link] [comments]
Related
- Is Deepseek V4 Pro working in Ollama?
- how to adjust the thinking effort for deepseek v4 on ollama cloud
- Ollama Cloud reliability + speed: 36-call bench across DeepSeek v3.2 → v4-pro → v4-flash + GLM-5.1
- Why Ollama Cloud doesn't have DeepSeek V4 Pro and Qwen3.6?
Source: r/ollama | 2026-07-25