Agents

Come check out GLM 5.1 in Code Arena for agentic web development tasks using tools. Don’t forget to vote, Code Arena scores are coming up ne…

GLM-5.1 is Zhipu AI's next-generation flagship model for agentic engineering, achieving state-of-the-art performance on SWE-Bench Pro and leading its predecessor GLM-5 by a wide margin on NL2Repo (...

DGX agentx-post
agentszhipu-ai--x

GLM-5.1 is Zhipu AI's next-generation flagship model for agentic engineering, achieving state-of-the-art performance on SWE-Bench Pro and leading its predecessor GLM-5 by a wide margin on NL2Repo (repo generation) and Terminal-Bench 2.0 (real-world terminal tasks). Code Arena (arena.ai/code) ranks AI models on agentic web development tasks using blind human evaluations, where users rate outputs without knowing which model produced them. GLM-5.1 ranked 3rd on Code Arena — ahead of GPT-5.4 and Gemini 3.1 Pro Preview — representing a +90-point jump over its predecessor GLM-5, making it the first open model to break into the top 3 on that leaderboard.

Related

Source: agents

Loading related sources…