Mistral OCR 3 launches at $2 per 1,000 pages with 74% win rate over previous version
Mistral AI released Mistral OCR 3, a document extraction model priced at $2 per 1,000 pages ($1 with Batch API discount). The model achieves a 74% overall win rate over its predecessor on forms, scanned documents, complex tables, and handwriting according to internal benchmarks.
Mistral OCR 3 launches at $2 per 1,000 pages with 74% win rate over previous version
Mistral AI released Mistral OCR 3 on December 17, 2024, a document extraction model priced at $2 per 1,000 pages, dropping to $1 per 1,000 pages with Batch API discount. The model is now available through API (identifier: mistral-ocr-2512) and via Document AI Playground in Mistral AI Studio.
Performance claims
According to Mistral AI, the model achieves a 74% overall win rate compared to Mistral OCR 2 across forms, scanned documents, complex tables, and handwriting. The company claims state-of-the-art accuracy compared to both enterprise document processing solutions and AI-native OCR solutions, though specific benchmark scores against named competitors were not disclosed.
Mistral evaluated the model using internal benchmarks based on customer use cases, comparing outputs to ground truth using fuzzy-match metrics for accuracy.
Technical capabilities
Mistral OCR 3 extracts text and embedded images from documents, outputting markdown enriched with HTML-based table reconstruction. The model handles:
- Handwriting: Cursive, mixed-content annotations, and handwritten text over printed forms
- Forms: Box detection, labels, handwritten entries, invoices, receipts, compliance forms
- Scanned documents: Compression artifacts, skew, distortion, low DPI, background noise
- Complex tables: Reconstructs structures with headers, merged cells, multi-row blocks, and column hierarchies using HTML table tags with colspan/rowspan
The model supports all languages and document form factors, representing what Mistral describes as a significant upgrade over OCR 2.
Pricing structure
- Standard API: $2 per 1,000 pages
- Batch API: $1 per 1,000 pages (50% discount)
- Annotations: $3 per 1,000 pages
- Self-hosting option available for organizations with data privacy requirements
The model is fully backward compatible with Mistral OCR 2.
What this means
At $2 per 1,000 pages standard pricing, Mistral OCR 3 undercuts typical enterprise document processing solutions that often charge per-page rates in the cents range. The 50% Batch API discount makes it particularly competitive for high-volume workflows. However, without public benchmark comparisons to models from Google, Amazon Textract, or other established OCR providers, customers will need to validate Mistral's performance claims in their own testing. The model's ability to handle multiple document types in a single solution could simplify document processing pipelines that currently require specialized tools for different formats.
Related Articles
Google DeepMind Launches Gemini 3.8 Live, Claims #1 Spot on Speech-to-Speech Benchmark
Google DeepMind has released Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, two voice-dialogue models that reason and execute background tasks without interrupting conversation. Google claims the Extended Thinking model ranks #1 on Artificial Analysis' Speech to Speech Quality Index with a score of 82.6.
Unverified 'GPT Astra' Model Appears on OpenRouter With 1.05M Token Context, No OpenAI Confirmation
OpenRouter is listing a model called 'OpenAI GPT Astra Latest' with a 1.05 million token context window and $10/$50 per-million-token pricing. OpenAI has made no public announcement, and the listing's own description says it is an auto-redirecting alias rather than a fixed model.
OpenRouter Lists 'GPT Sol Latest' — An Alias Pointer to OpenAI's Newest Sol-Family Model, Not a Standalone Release
OpenRouter has added a listing called '~openai/gpt-sol-latest,' described as an alias that always points to the newest model in an undisclosed 'GPT Sol' family from OpenAI. The listing shows a 1050K token context window and pricing of $2.00 per million input tokens and $10.00 per million output tokens, but OpenAI has not publicly confirmed a model line by this name.
OpenAI Launches GPT-Live-1 API for Full-Duplex Voice Apps That Talk and Listen Simultaneously
OpenAI has released GPT-Live-1 as a developer API, a speech model capable of full-duplex conversation—listening and talking simultaneously. It already powers ChatGPT's voice mode and costs $0.05 per minute, with benchmark scores showing sharp improvements over GPT-Realtime-2.1.
Comments
Loading...