Mistral OCR 3 launches at $2 per 1,000 pages with 74% win rate over previous version
Mistral AI released Mistral OCR 3, a document extraction model priced at $2 per 1,000 pages ($1 with Batch API discount). The model achieves a 74% overall win rate over its predecessor on forms, scanned documents, complex tables, and handwriting according to internal benchmarks.
Mistral OCR 3 launches at $2 per 1,000 pages with 74% win rate over previous version
Mistral AI released Mistral OCR 3 on December 17, 2024, a document extraction model priced at $2 per 1,000 pages, dropping to $1 per 1,000 pages with Batch API discount. The model is now available through API (identifier: mistral-ocr-2512) and via Document AI Playground in Mistral AI Studio.
Performance claims
According to Mistral AI, the model achieves a 74% overall win rate compared to Mistral OCR 2 across forms, scanned documents, complex tables, and handwriting. The company claims state-of-the-art accuracy compared to both enterprise document processing solutions and AI-native OCR solutions, though specific benchmark scores against named competitors were not disclosed.
Mistral evaluated the model using internal benchmarks based on customer use cases, comparing outputs to ground truth using fuzzy-match metrics for accuracy.
Technical capabilities
Mistral OCR 3 extracts text and embedded images from documents, outputting markdown enriched with HTML-based table reconstruction. The model handles:
- Handwriting: Cursive, mixed-content annotations, and handwritten text over printed forms
- Forms: Box detection, labels, handwritten entries, invoices, receipts, compliance forms
- Scanned documents: Compression artifacts, skew, distortion, low DPI, background noise
- Complex tables: Reconstructs structures with headers, merged cells, multi-row blocks, and column hierarchies using HTML table tags with colspan/rowspan
The model supports all languages and document form factors, representing what Mistral describes as a significant upgrade over OCR 2.
Pricing structure
- Standard API: $2 per 1,000 pages
- Batch API: $1 per 1,000 pages (50% discount)
- Annotations: $3 per 1,000 pages
- Self-hosting option available for organizations with data privacy requirements
The model is fully backward compatible with Mistral OCR 2.
What this means
At $2 per 1,000 pages standard pricing, Mistral OCR 3 undercuts typical enterprise document processing solutions that often charge per-page rates in the cents range. The 50% Batch API discount makes it particularly competitive for high-volume workflows. However, without public benchmark comparisons to models from Google, Amazon Textract, or other established OCR providers, customers will need to validate Mistral's performance claims in their own testing. The model's ability to handle multiple document types in a single solution could simplify document processing pipelines that currently require specialized tools for different formats.
Related Articles
Moonshot AI releases 2.8T parameter Kimi K3, pricing at $3/$15 per million tokens
Chinese AI lab Moonshot AI released Kimi K3, a 2.8 trillion parameter model priced at $3 per million input tokens and $15 per million output tokens. The model is currently available via API, with open weights promised by July 27, 2026. This represents the most expensive pricing from a Chinese AI lab to date, matching Anthropic's Claude Sonnet series.
Thinking Machines Lab releases Inkling: 975B-parameter open-weights multimodal model under Apache-2.0
Thinking Machines Lab released Inkling, a Mixture-of-Experts transformer with 975B total parameters and 41B active parameters, trained on 45 trillion tokens of text, images, audio and video. The Apache-2.0 licensed model is designed as a base for fine-tuning rather than a frontier model.
Moonshot AI Releases Kimi K3: Open-Weight Multimodal Reasoning Model with 1M Context Window
Moonshot AI has released Kimi K3, an open-weight multimodal reasoning model with a 1-million token context window. The model is priced at $3 per 1M input tokens and $15 per 1M output tokens, available through OpenRouter.
Google releases Gemma 4 E2B, optimized to run natively on Pixel 10's Tensor G5 TPU
Google has released Gemma 4 E2B for TPU, a variant of its open-source Gemma 4 model optimized to run natively on the Tensor G5 chip in Pixel 10 devices. The multimodal model enables completely offline AI chat, image recognition, and audio transcription on Pixel 10, 10 Pro, 10 Pro XL, and 10 Pro Fold.
Comments
Loading...