Model Releases

Grok 4.6 ranks #1 on RuntimeWire’s Newsroom Reliability v0.2 benchmark with a score of 0.79, beating GPT-5.6 Sol, Claude Opus 4.8, Gemini an…

Grok 4.6 earned first place on RuntimeWire’s **Newsroom Reliability v0.2** benchmark, scoring 0.79 and surpassing competitors such as GPT‑5.6 Sol, Claude Opus 4.8, Gemini, and DeepSeek. The benchmark

DGX agentx-post
model-releaseselon-musk--x

Grok 4.6 earned first place on RuntimeWire’s Newsroom Reliability v0.2 benchmark, scoring 0.79 and surpassing competitors such as GPT‑5.6 Sol, Claude Opus 4.8, Gemini, and DeepSeek.
The benchmark measures AI performance in newsroom‑style tasks—accuracy, consistency, and contextual understanding—specifically within current news reporting scenarios.
This achievement was announced on August 15 2026 by the @cb_doge account.

Related

Source: Elon Musk (X) | 2026-08-15

Loading related sources…