Model Releases

We benchmarked Opus 5 comprehensively on document understanding through ParseBench. It is roughly on par with Opus 4.8 - it does a few point…

We benchmarked Opus 5 comprehensively on document understanding through ParseBench. It is roughly on par with Opus 4.8 - it does a few points worse on dense tables, but does slightly better on parsing

DGX agentx-post
model-releasesjerry-liu--x

We benchmarked Opus 5 comprehensively on document understanding through ParseBench. It is roughly on par with Opus 4.8 - it does a few points worse on dense tables, but does slightly better on parsing charts and visual grounding. For reference, Gemini 3.6 Flash has better results on tables, is slightly worse on charts, and is half the price. LlamaParse agentic is better on all fronts (including tables) at 1/6th of the price. tl;dr use Opus 5 as much as you want for coding and knowledge work. But at 8c per page and average OCR performance, don't use it to parse documents at scale Our full ParseBench leaderboard which we periodically update is here: https://www.parsebench.ai/

Related

Source: Jerry Liu (X) | 2026-07-24

Loading related sources…