3 posts tagged ocr.
OpenAI builds its own inference chip, Mistral and Baidu both ship serious OCR upgrades, and Anthropic's Claude is writing 65% of its own team's code.
Three serious OCR and document extraction models dropped in the same week, and the gap between 'parsing a PDF' and 'understanding a document' quietly closed.
A compact 0.9B multimodal model that handles document parsing, tables, and structured extraction without burning through compute resources.