Model Releases
We benchmarked Gemini 3.6 Flash and Gemini 3.5 Flash Lite on document understanding. We compared against their prior versions - Gemini 3.5 F…
We benchmarked Gemini 3.6 Flash and Gemini 3.5 Flash Lite on document understanding. We compared against their prior versions - Gemini 3.5 Flash and Gemini 3.1 Flash Lite. 1️⃣ Gemini 3.6 Flash has rou
We benchmarked Gemini 3.6 Flash and Gemini 3.5 Flash Lite on document understanding. We compared against their prior versions - Gemini 3.5 Flash and Gemini 3.1 Flash Lite. 1️⃣ Gemini 3.6 Flash has roughly similar performance, though it does go down 14% in chart understanding 2️⃣ Gemini 3.5 Flash Lite improves on layout detection by 11%, though it regresses on tables by ~12% In general it seems like while Gemini 3 Flash was well tuned for document understanding, subsequent versions have been posttrained for coding and reasoning, which has led to a plateau in visual recognition capabilities. It would be interesting to see if Google starts to prioritize visual understanding again with the Flash series, or if they're also going all in on reasoning models. Check out these results and more on ParseBench: https://www.parsebench.ai/
Related
- We pit LlamaParse against frontier models (Opus 4.6, Gemini 3.1 Pro, GPT-5.4) in a live OCR arena. ICYMI: the full workshop is on Youtube! F…
- Document Parsing + Gemini 🔥 Excited to collaborate with the Google team on this, here's to many more!
- Want to see which frontier models do the best on document understanding? Check out our ParseBench leaderboard on @kaggle! https://www.kaggle…
- If you want to stack rank LLMs/VLMs on document understanding 📄, you can through ParseBench, now live on @kaggle 📊 ParseBench is the most …
Source: Jerry Liu (X) | 2026-07-22