We need more evals for document understanding. ParseBench is a really great start. I respect @llama_index ‘s work on this. 📈
We need more evals for document understanding. ParseBench is a really great start. I respect @llama_index ‘s work on this. 📈 We benchmarked GPT-5.5 on document understanding 📄📊 We ran it through Parse