Model Releases
We are rolling out GLM-5.3 on Ollama. Private. Fast. US and Europe hosted. No data retention. Try Ollama's GLM-5.3-Flash as we bring it GLM-…
We are rolling out GLM-5.3 on Ollama. Private. Fast. US and Europe hosted. No data retention. Try Ollama's GLM-5.3-Flash as we bring it GLM-5.3 online. Super fast. Claude Code: ollama launch claude --
We are rolling out GLM-5.3 on Ollama. Private. Fast. US and Europe hosted. No data retention. Try Ollama's GLM-5.3-Flash as we bring it GLM-5.3 online. Super fast. Claude Code: ollama launch claude --model glm-5.3-flash:cloud OpenCode: ollama launch opencode --model glm-5.3-flash:cloud Hermes Agent: ollama launch hermes --model glm-5.3-flash:cloud or plug it into any other harness/apps. API endpoint is also available. GLM-5.3 is now open-weight. Our most capable model for agentic coding and cyber defense is now available to download, run, and customize. Weights: https://huggingface.co/zai-org/GLM-5.3 Tech blog: https://z.ai/blog/glm-5.3
Related
- We just added significantly more NVIDIA Blackwell GPUs to better serve GLM-5.1 model on Ollama's cloud. We have been adding more GPUs daily …
- ollama launch claude --model glm-5.1:cloud (We are rushing to get more capacity 🙏🙏🙏)
- Currently serving 200 tps+ output speed for DeepSeek-V4-Flash on Ollama's cloud with zero data retention (ZDR). Have an amazing weekend 🫡
Source: Ollama (X) | 2026-08-28