Applications
Grok continues to lead global benchmarks: • #1 in AA Omniscience (lowest hallucination rate) • #1 in IFBench performance • #1 on BridgeBench…
Grok continues to lead global benchmarks: • #1 in AA Omniscience (lowest hallucination rate) • #1 in IFBench performance • #1 on BridgeBench Reasoning • #1 on BridgeBench Speed • #1 on BridgeBench Low
Grok continues to lead global benchmarks: • #1 in AA Omniscience (lowest hallucination rate) • #1 in IFBench performance • #1 on BridgeBench Reasoning • #1 on BridgeBench Speed • #1 on BridgeBench Lowest Hallucination • #1 in Text Arena (medicine & healthcare) • #1 in Video Edit Arena • #1 in AlphaArena Leaderboard Grok is dominating. 🔥
Related
- Grok
- There are more competitive small model makers, but there is still a very big gap between what small models can do and what large models can …
- It isn't at the level of the Big Three models when you poke at it, but a very solid start.
- Anyhow, its not bad. Just not the vibe level that the benchmarks might indicate. And, for a first re-entry into the frontier model space, gi…
Source: Elon Musk (X) | 2026-04-13