Mistral OCR (25.05)

Mistral OCR (25.05) is an Optical Character Recognition API for document understanding. Mistral OCR (25.05) excels in understanding complex document elements, including interleaved imagery, mathematical expressions, tables, and advanced layouts such as LaTeX formatting. The model enables deeper understanding of rich documents such as scientific papers with charts, graphs, equations and figures.

Mistral OCR (25.05) is an ideal model to use in combination with a RAG system that takes multimodal documents (such as slides or complex PDFs) as input.

You can couple Mistral OCR (25.05) with other Mistral models to reformat the results. This combination ensures that the extracted content is not only accurate but also presented in a structured and coherent manner, making it suitable for various downstream applications and analyses.

View model card in Model Garden

Model ID mistral-ocr-2505
Launch stage GA
Supported inputs & outputs
  • Inputs:
    Documents
  • Outputs:
    Text
Usage types
Versions
  • Mistral OCR (25.05)
    • Launch stage: GA
    • Release date: May 14, 2025
Supported regions

Model availability

  • United States
    • us-central1
  • Europe
    • europe-west4

ML processing

  • United States
    • Multi-region
  • Europe
    • Multi-region
Quota limits

us-central1:

  • QPM: 30
  • Pages per request: 30 (1 page = 1 million input tokens and 1 million output tokens)
  • Context length: 30 pages
  • Max request size: 30MB for streaming, 10MB for unary

europe-west4:

  • QPM: 30
  • Pages per request: 30 (1 page = 1 million input tokens and 1 million output tokens)
  • Context length: 30 pages
  • Max request size: 30MB for streaming, 10MB for unary

Pricing See Pricing.