Model Releases
NuExtract3 released: open-weight 4B VLM for Markdown, OCR and structured extraction (self-hostable) [P]
NuExtract3 is a unified 4B vision-language reasoning model for document understanding that combines structured information extraction with image-to-Markdown conversion, suitable for OCR and RAG prepro
NuExtract3 is a unified 4B vision-language reasoning model for document understanding that combines structured information extraction with image-to-Markdown conversion, suitable for OCR and RAG preprocessing across documents like scans, receipts, forms, and invoices. The model supports both reasoning and non-reasoning inference modes , and is available as open-weight for self-hosting deployment.
Source: r/MachineLearning | 2026-05-22