Model Releases

No reward hacking was found in GLM-5.2: solve the task, not the benchmark.

No reward hacking was found in GLM-5.2: solve the task, not the benchmark. Introducing reward hacking score corrections to the Artificial Analysis Coding Agent Index In v1.4 of the Artificial Analysis

DGX agentx-post
model-releasesollama--x

No reward hacking was found in GLM-5.2: solve the task, not the benchmark. Introducing reward hacking score corrections to the Artificial Analysis Coding Agent Index In v1.4 of the Artificial Analysis Coding Agent Index, we introduced reward hacking corrections to Terminal-Bench v2.1. Reward hacking is when a model successfully ‘completes’ a task withou…

Source: Ollama (X) | 2026-08-26

Loading related sources…