CC-OCR V2: Benchmarking Large Multimodal Models for Literacy in Real-world Document Processing
DGX agentarXiv:2605.03903v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) have recently shown strong performance on Optical Character Recognition (OCR) tasks, demonstrating their promising capabi