
Mistral's new OCR model beats competitors in 72 percent of blind test cases, company says
Quick Answer
Mistral AI's new OCR 4 model outperforms competitors in 72% of blind tests, showcasing its superior text recognition capabilities across various document formats.
Quick Take
This advancement positions Mistral as a leader in the OCR space, particularly for users needing accurate document processing from PDFs, Word files, and PowerPoint presentations.
Key Points
- OCR 4 model reads text from PDFs, Word files, and PowerPoint presentations.
- Achieved 72% success rate in blind test comparisons against competitors.
- Mistral AI strengthens its position in the OCR technology market.
- Improved accuracy benefits users requiring reliable document processing.
- The model's performance could influence future OCR developments.
📖 Reader Mode
~1 min readMistral AI has released OCR 4, a new model that reads text from documents like PDFs, Word files, and PowerPoint presentations.
Unlike earlier versions, OCR 4 doesn't just pull raw text. It also identifies where each element sits on the page and what role it plays - whether it's a title, a table, an equation, or a signature. This block classification helps break documents into meaningful sections automatically, useful for feeding them into search systems or letting AI agents process them. The model also outputs confidence scores, giving an estimate of how certain it is about each word or page it reads.

OCR 4 supports 170 languages and works well even with less common ones, according to Mistral. In a blind test with over 600 documents, independent reviewers preferred OCR 4's results 72 percent of the time over competing models, the company says. The model is available through the API, Mistral Studio, and Microsoft Foundry. It costs $4 per 1,000 pages, or $2 in batch mode.
— Originally published at the-decoder.com
Want this in your inbox every morning?
Daily brief at your local 8am — bilingual EN/中文, free.
More from The Decoder
See more →
An AI model programmed nonstop for 19 days on a single MirrorCode task that cost $2,600 to run
Epoch AI's MirrorCode benchmark reveals Claude Opus 4.7 as the leader with a 56% solve rate, reconstructing a 16,000-line toolkit in 14 hours. Despite this, all models tested struggle with the most complex tasks, highlighting limitations in current AI capabilities. The single task consumed $2,600 over 19 days, raising questions about cost-effectiveness in AI development.

