Agents

We built one of the most comprehensive benchmarks for document extraction, and evaluated it across a lot of different systems: ✅ one-shot fr…

We built one of the most comprehensive benchmarks for document extraction, and evaluated it across a lot of different systems: ✅ one-shot frontier VLMs ✅ frontier VLMs + coding agent harnesses ✅ one-s

DGX agentx-post
agentsjerry-liu--x

We built one of the most comprehensive benchmarks for document extraction, and evaluated it across a lot of different systems: ✅ one-shot frontier VLMs ✅ frontier VLMs + coding agent harnesses ✅ one-shot open weight VLMs ✅ other document extraction tools (including LlamaParse) Document extraction is an extremely diverse task that covers many different types of docs. From simple schemas over short docs (e.g. resume extraction) to complex extraction docs (credit agreements, data room bundles). @disiok is leading this webinar. Come check it out! https://watch.getcontrast.io/register/llamaindex-inside-extractbench-benchmarking-document-extraction-for-agents Every extraction API demos well on a clean invoice. But what about the scanned form, the nested table, the 40-page financial report with merged headers? We tested 14 frontier systems to find out. ExtractBench evaluates schema-guided extraction across 370 enterprise documents, 67 …

Related

Source: Jerry Liu (X) | 2026-08-26

Loading related sources…