Model Releases
Possible memory leak in Ollama when using Claude Code?
Users in the r/ollama community have reported a possible memory leak occurring in Ollama when it is used as a backend with Claude Code, with Ollama runner processes not always being properly termin...
Users in the r/ollama community have reported a possible memory leak occurring in Ollama when it is used as a backend with Claude Code, with Ollama runner processes not always being properly terminated even after models are unloaded. Up to version v0.7.0, Ollama is known to suffer from a sporadic VRAM memory leak that gradually fills GPU memory over time, potentially degrading system performance or preventing new models from loading. Workarounds discussed include scripted detection tools that can identify lingering runner processes and safely restart the Ollama service to reclaim memory.
Related
- Claude Code Cheat Sheet Here is a list of top commands, shortcuts, and patterns you need to know when using Claude Code. We will update this…
- Enjoy faster @-mentions in Claude Code! This went out in v2.1.85.
- 'Claude Code isn't magic. The harness layer is just software, and software is something any dev can shape to fit how they want to work.' Che…
- There were some exceptionally cool demos from @ollama and omlx using MLX to run Qwen 3.5 and Gemma 4 on Apple silicon. The capabilities of l…
- Google's Gemma 4 is pretty wild. You can now run it locally with OpenClaw in 3 steps. 1. Install Ollama 2. Pull Gemma 4 model 3. Launch Open…
Source: model-releases