Local Ai

How much is the GPU usage?

I am trying to decide on buying the ollama pro subscription. But their usage policy is vague as hell. I don't mind the 'GPU usage time' but how much GPU time do I actually get? I still don't seem to f

DGX agentreddit
local-air-ollama

I am trying to decide on buying the ollama pro subscription. But their usage policy is vague as hell. I don't mind the "GPU usage time" but how much GPU time do I actually get? I still don't seem to find any hard data like how much token you can squeeze out of on a single 5 hr session on average with a fixed model. The stuff is so vague that ollama cloud could randomly half the usage limits without the consumer having no way to know or at least proof. If someone has actually done any experiment on the limits or collected hard data it would be immensely helpful to the community. For example anything like if it's possible to generate 10k input and 1k output with 95% cache hit on 5 hours session using model X. I am simply trying to make an informed decisions as there are a lot of providers that offer better deal than API pricing. submitted by /u/sakibshahon [link] [comments]

Related

Source: r/ollama | 2026-07-26

Loading related sources…