Model Releases

INCREDIBLE GLM-5.1 weights are now opensource > i’ve had early access to the weights for the past few days > and yeah… this one matters a lo…

INCREDIBLE GLM-5.1 weights are now opensource > i’ve had early access to the weights for the past few days > and yeah… this one matters a lot benchmarks? > SWE-Bench Pro: 58.4 > beats Opus 4.6 (57.3)

DGX agentx-post
model-releaseszhipu-ai--x

INCREDIBLE GLM-5.1 weights are now opensource > i’ve had early access to the weights for the past few days > and yeah… this one matters a lot benchmarks? > SWE-Bench Pro: 58.4 > beats Opus 4.6 (57.3) > beats GPT-5.4 (57.7) > beats Gemini 3.1 Pro (54.2) let that sink in open weights beating closed > open-weight (MIT licensed) > built for agentic engineering > sustains long-horizon reasoning > runs locally via vLLM / SGLang / Transformers but the real unlock isn’t just first-pass scores it’s this: > a year ago, agents used to do ~20 steps > GLM-5.1 can do ~1,700 steps > longer runs > more iteration > better results over time and now you can verify it yourself, locally > the gap between opensource and closed models? still ~6 months or less and shrinking fast Introducing GLM-5.1: The Next Level of Open Source - Top-Tier Performance: #1 in open source and #3 globally across SWE-Bench Pro, Terminal-Bench, and NL2Repo. - Built for Long-Horizon Tasks: Runs autonomously for 8 hours, refining strategies through thousands of iterations. Blog…

Related

Source: Zhipu AI (X) | 2026-04-07

Loading related sources…