Tried running LLMs locally to save API costs… ended up waiting 13 minutes for ONE response 🤡
A Reddit post in r/ollama describes a user's experience attempting to run LLMs locally via Ollama to avoid cloud API costs, only to encounter severely degraded performance — waiting 13 minutes for ...