Local Ai
How much is the GPU usage?
I am trying to decide on buying the ollama pro subscription. But their usage policy is vague as hell. I don't mind the 'GPU usage time' but how much GPU time do I actually get? I still don't seem to f
I am trying to decide on buying the ollama pro subscription. But their usage policy is vague as hell. I don't mind the "GPU usage time" but how much GPU time do I actually get? I still don't seem to find any hard data like how much token you can squeeze out of on a single 5 hr session on average with a fixed model. The stuff is so vague that ollama cloud could randomly half the usage limits without the consumer having no way to know or at least proof. If someone has actually done any experiment on the limits or collected hard data it would be immensely helpful to the community. For example anything like if it's possible to generate 10k input and 1k output with 95% cache hit on 5 hours session using model X. I am simply trying to make an informed decisions as there are a lot of providers that offer better deal than API pricing. submitted by /u/sakibshahon [link] [comments]
Related
- How much usage does Ollama Pro give right now vs direct API?
- Ollama has reduced the limits on their Pro subscription.
- Trying to understand the Ollama debate. What’s actually going on?
- Another example of greed. The PRO subscription!
- How are you guys actually burning through your Pro quota so fast?
Source: r/ollama | 2026-07-26