Model Releases
Good reason to try Grok 4.5 with Grok Build. It gets better every day!
Good reason to try Grok 4.5 with Grok Build. It gets better every day! Grok 4.5 just took the #1 spot on the Long-Horizon Terminal-Bench, outperforming Claude Fable 5, Claude Opus 4.8 and GPT-5.6-sol
Good reason to try Grok 4.5 with Grok Build. It gets better every day! Grok 4.5 just took the #1 spot on the Long-Horizon Terminal-Bench, outperforming Claude Fable 5, Claude Opus 4.8 and GPT-5.6-sol This benchmark tests whether an AI agent can sustain progress across hundreds of dependent terminal actions for up to 90 minutes without losing the thr…
Related
- Grok Build
- SpaceXAI's Grok 4.5 takes the #1 spot on AutomationBench-AA with a score of 51%, ahead of Claude Fable 5 (49%) and Claude Opus 4.8 (48%) at …
- Grok 4.5 is dominating the latest AI leaderboards Claims the #1 spot: • #1 on AutomationBench-AA • #1 on Terminal-Bench v2 • #1 on Harvey Le…
- Grok 4.5 in Grok Build also stands out for its efficiency. Grok 4.5 in Grok Build cost 2.49 per task while Fable 5 in Claude Code cost 11.…
Source: Elon Musk (X) | 2026-07-15