Agents
Come check out GLM 5.1 in Code Arena for agentic web development tasks using tools. Don’t forget to vote, Code Arena scores are coming up ne…
GLM-5.1 is Zhipu AI's next-generation flagship model for agentic engineering, achieving state-of-the-art performance on SWE-Bench Pro and leading its predecessor GLM-5 by a wide margin on NL2Repo (...
GLM-5.1 is Zhipu AI's next-generation flagship model for agentic engineering, achieving state-of-the-art performance on SWE-Bench Pro and leading its predecessor GLM-5 by a wide margin on NL2Repo (repo generation) and Terminal-Bench 2.0 (real-world terminal tasks). Code Arena (arena.ai/code) ranks AI models on agentic web development tasks using blind human evaluations, where users rate outputs without knowing which model produced them. GLM-5.1 ranked 3rd on Code Arena — ahead of GPT-5.4 and Gemini 3.1 Pro Preview — representing a +90-point jump over its predecessor GLM-5, making it the first open model to break into the top 3 on that leaderboard.
Related
- GLM-5.1 is now available in Windsurf! Try it out and let us know what you think
- GLM-5.1 is live everywhere you use the Kilo Gateway (VS Code extension, Cloud Agents, KiloClaw, etc). Thank you @Zai_org! ⚡️
- GLM-5.1 gives teams a stronger model for coding, tool use, and sustained agent performance on Together AI. Learn more: http://www.together.a…
- Introducing GLM-5.1 from @Zai_org on Together AI. AI natives can now use GLM-5.1 on Together and benefit from reliable inference for product…
- 智谱直接把开源 Agent 拉到新高度了! GLM-5.1 正式开源: ✅ 开源模型里 SWE-Bench Pro 拿下 #1(58.4),全球第 3 ✅ 真正长时程 Agent:自主运行 8 小时、几千次迭代 + 自审循环 ✅ 从零搭出一个带 50+ App 的完整 Linux…
- 🤖 【2026最新】1秒で理解!8時間自律稼働する「GLM-5.1」がLinuxをゼロから構築。オープンソースが世界を制す極限全貌 【2026最新】1秒で理解!8時間自律稼働し、Linuxデスクトップをゼロから完成させる衝撃のオープンソースAI「GLM-5.1」が登場。SWE-…
Source: agents