Model Releases
A lot of you are asking about the :cloud tag. Here's the deal: GLM-5.1 is a 744B parameter beast. To hit that 54.9 benchmark score without t…
A lot of you are asking about the :cloud tag. Here's the deal: GLM-5.1 is a 744B parameter beast. To hit that 54.9 benchmark score without turning my M4 Mac Mini into a space heater, I run the cloud e
A lot of you are asking about the :cloud tag. Here's the deal: GLM-5.1 is a 744B parameter beast. To hit that 54.9 benchmark score without turning my M4 Mac Mini into a space heater, I run the cloud endpoint to push the model to experiment. (But you have to pay Ollama) If you've got a monster server rack and want to go fully local, here's your command: ollama run glm-5.1 TL;DR: Cloud = build fast. Local = you're a hardware god.
Related
- To run GLM-5.1 locally (744B params, 40B active MoE), full precision needs ~1.65TB disk + enterprise hardware like 8x H200/B200 GPUs. Minimu…
- The chart says GLM-5.1 scored 54.9 on coding benchmarks. Three points behind Claude Opus 4.6. Interesting but not the story. The story is wh…
- GLM 5.1 is now LIVE in Atomic Chat SOTA for code & chat – now runs locally with TurboQuant Thanks to @zai_org for open-sourcing this frontie…
- INCREDIBLE GLM-5.1 weights are now opensource > i’ve had early access to the weights for the past few days > and yeah… this one matters a lo…
- Check out the GLM-5.1 first impressions with Peter on our YouTube https://www.youtube.com/watch?v=f11tVBXWr2g
- GLM-5.1 is live everywhere you use the Kilo Gateway (VS Code extension, Cloud Agents, KiloClaw, etc). Thank you @Zai_org! ⚡️
Source: Zhipu AI (X) | 2026-04-07