Back to News
Mistral AIJune 23, 2026

Mistral OCR 4

Paid
CV
ML
Testing

Explore Mistral OCR 4

Visit the official website to learn more and get started

### TL;DR

Mistral OCR 4 is an advanced Optical Character Recognition (OCR) model designed to extract text and structured content from documents with high accuracy. It offers features like bounding boxes, block classification, and inline confidence scores, supporting 170 languages across 10 language groups. The model is available for self-hosted deployments and integrates seamlessly with Mistral's Search Toolkit for enterprise search and retrieval-augmented generation (RAG) pipelines.

Key Insights & Metrics

Pricing
OCR — $4/ 1,000 pages; Batch-API — $2 / 1,000 pages; Document AI — $5 / 1000 pages
Cost structure
Version
4
Current release version
Hardware
No specific hardware requirements
Compute requirements
Category
Paid
Licensing model
Region
France
Primary region

Key Features

  • Bounding boxes for text localization
  • Block classification (titles, tables, equations, signatures)
  • Inline confidence scores for extracted text
  • Support for 170 languages across 10 language groups
  • Integration with Mistral Search Toolkit for enterprise search and RAG

Related Releases

Step-Audio-R1

Step-Audio-R1 is an advanced audio language model developed by StepFun AI, designed to enhance audio reasoning capabilities by grounding its reasoning in acoustic features. It introduces Modality-Grounded Reasoning Distillation (MGRD), an iterative training framework that shifts the model's reasoning from textual abstractions to acoustic properties, effectively addressing the 'inverted scaling' problem where performance degrades with longer reasoning. This model has demonstrated superior performance across various audio understanding and reasoning benchmarks, surpassing models like Gemini 2.5 Pro and achieving results comparable to Gemini 3 Pro.

StepFun AINov 29
Open

SkyRL tx

SkyRL tx is an open-source library that implements a backend for the Tinker API, enabling users to set up their own Tinker-like services on personal hardware. It supports end-to-end reinforcement learning (RL) and offers significantly faster sampling. The library is designed to be modular, allowing easy prototyping of new training algorithms, environments, and execution plans without compromising usability or speed.

NovaSky AINov 3
Open

Z-Image

Z-Image is an efficient image generation foundation model developed by Tongyi-MAI, designed to produce high-quality, diverse, and stylistically versatile images. It serves as a robust base for creators, researchers, and developers seeking advanced image generation capabilities.

Tongyi-MAINov 27
Open

SINQ

SINQ (Sinkhorn-Normalized Quantization) is a novel, fast, and high-quality quantization method designed to make any Large Language Model (LLM) smaller while preserving accuracy. It offers a plug-and-play, model-agnostic technique that delivers state-of-the-art performance for LLMs without sacrificing accuracy.

HuaweiNov 14
Open

Discussion

0
Upvotes
0
Downvotes
0 reviews

Sign in to leave a review

Reviews

No reviews yet. Be the first to review!

🚀 Join the AI dev community — follow us everywhere

© 2026 MARKTECHPOST AI MEDIA INC. All rights reserved.Terms & ConditionsPrivacy Policy
Beta Mode