Model Releases
Ollama + llama.cpp + VS Code + OWUI + ...cloud?
I've been trying to figure out what the best configuration for my local stack would be and have been primarily determining this through trial and error. I wanted to get some outside opinions and stack
I've been trying to figure out what the best configuration for my local stack would be and have been primarily determining this through trial and error. I wanted to get some outside opinions and stack recommendations to kind of broaden my knowledge in this arena. I've got Ollama as the base of everything I run for AI inference and love it. I don't have the compute to run any meaningfully sized models on my hardware and was reading up a little on the cloud options last night and would really appreciate some feedback from people that have used it. How are you running it? What does your stack look like? What do you like about it? Thanks, guys! submitted by /u/SkootinSkitzo [link] [comments]
Related
- What's your local AI coding setup on a MacBook Pro M4?
- I have a Macbook AIR M5 Base and I want to run an Agentic Coding program, similar to Claude Code or Codex. Besides the model, how do I do it? I've already tried with Ollama, VS Code, Opencode, and haven't been able to. (I'm not a developer, sorry)
- No Claude sub, and I'm still getting the 'Claude Code experience' by routing OpenAI, Copilot, GLM, and Ollama through one harness
- I built a VS Code extension that cuts my Claude API bill to ~$5/day
Source: r/ollama | 2026-08-22