Model Releases
We tuned an AI agent that can do large-scale document extraction from long docs (50+ pages, some with 10k-100k fields) with 94%+ accuracy 📈…
We tuned an AI agent that can do large-scale document extraction from long docs (50+ pages, some with 10k-100k fields) with 94%+ accuracy 📈 It uses a harness + model set that is tuned specifically for
We tuned an AI agent that can do large-scale document extraction from long docs (50+ pages, some with 10k-100k fields) with 94%+ accuracy 📈 It uses a harness + model set that is tuned specifically for reasoning over extracting out complex information from complex docs. Each extracted field comes with a confidence score as well as a bounding box denoting where it came from. It does 10-20% better in accuracy than generalized coding agent harnesses (e.g. Claude Code Opus 4.8 and Codex GPT-5.6). Check out the video below for a demo. The mode is called LlamaExtract Agentic Plus. Learn more about our extraction benchmark and agentic plus mode here: https://www.llamaindex.ai/blog/introducing-extractbench If you have very complex document extraction needs, come check out LlamaParse! https://cloud.llamaindex.ai/ Media Our "agentic plus" extractor in LlamaParse is great for extracting out massive volumes of fields (e.g. 10k-100k+ fields) from long documents (100-500 pages) We tested with a doc that contains a giant matrix of all creditors for FTX 💸 (75k fields, 114 pages). See screenshot below…
Related
- This week we launched the world's most accurate document extraction agent over real-world documents. Introducing LlamaExtract Agentic Plus …
- ExtractBench is one of the most comprehensive benchmarks for real-world document extraction. ✅ It covers 4869 pages, across 67 document type…
- Our 'agentic plus' extractor in LlamaParse is great for extracting out massive volumes of fields (e.g. 10k-100k+ fields) from long documents…
Source: Jerry Liu (X) | 2026-08-16