Provider

Mistral AI

Every Mistral AI model UltraGPT can reach, kept in sync with the live catalog — pricing, context window and capabilities for each one.

25 models from Mistral AI Up to 262.1K tokens of context
Open the appBrowse every provider

25

Models tracked

1

Providers

4

Reasoning models

15

See images (vision)

262K

Largest context window

0

Free to run

Directory

All Mistral AI models.

Search and filter this provider's lineup — pricing per million tokens, context window, modalities, reasoning effort and independent benchmark scores where available.

25 of 25 models

This is Mistral AI's flagship model, Mistral Large 2 (version `mistral-large-2407`). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more. Read the launch announcement [here](https://mistral.ai/news/mistral-large-2407/)....

Tool callingStreamingJSON mode
Context window128K tok
Max output102.4K tok
ReleasedFeb 26, 2024
Knowledge cutoffNov 30, 2024
Text + FileText
Input$2 / 1M tok
Output$6 / 1M tok
Cache read$0.2 / 1M tok

This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more. Read the launch announcement [here](https://mistral.ai/news/mistral-large-2407/)....

Tool callingStreamingJSON mode
Context window131.1K tok
Max output104.9K tok
ReleasedNov 19, 2024
Knowledge cutoffMar 31, 2024
Text + FileText
Input$2 / 1M tok
Output$6 / 1M tok
Cache read$0.2 / 1M tok

Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)

Tool callingStreamingJSON mode
Context window256K tok
Max output204.8K tok
ReleasedAug 1, 2025
Knowledge cutoffMar 31, 2025
Text + FileText
Input$0.3 / 1M tok
Output$0.9 / 1M tok
Cache read$0.03 / 1M tok

Design Arena — ranked #97 across 6 categories

Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)

Tool callingStreamingJSON mode
Context window256K tok
Max output204.8K tok
ReleasedAug 1, 2025
Knowledge cutoffMar 31, 2025
Text + FileText
Input$0.15 / 1M tok
Output$0.45 / 1M tok
Cache read$0.015 / 1M tok

Design Arena — ranked #97 across 6 categories

Devstral 2 is a state-of-the-art open-source model by Mistral AI specializing in agentic coding. It is a 123B-parameter dense transformer model supporting a 256K context window. Devstral 2 supports exploring...

Tool callingStreamingJSON mode
Context window262.1K tok
Max output209.7K tok
ReleasedDec 9, 2025
Knowledge cutoffDec 1, 2025
Text + FileText
Input$0.4 / 1M tok
Output$2 / 1M tok
Cache read$0.04 / 1M tok
Intelligence9.4
Coding31.3
Agentic4.9

The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B counterpart. A powerful and efficient language...

Tool callingVisionStreamingJSON mode
Context window262.1K tok
Max output209.7K tok
ReleasedDec 2, 2025
Knowledge cutoff
Text + ImageText
Input$0.2 / 1M tok
Output$0.2 / 1M tok
Cache read$0.02 / 1M tok
Intelligence6.0
Coding14.4
Agentic1.1

Design Arena — ranked #98 across 4 categories

The smallest model in the Ministral 3 family, Ministral 3 3B is a powerful, efficient tiny language model with vision capabilities.

Tool callingVisionStreamingJSON mode
Context window131.1K tok
Max output104.9K tok
ReleasedDec 2, 2025
Knowledge cutoff
Text + ImageText
Input$0.1 / 1M tok
Output$0.1 / 1M tok
Cache read$0.01 / 1M tok
Intelligence4.8
Coding4.8
Agentic0.8

Design Arena — ranked #107 across 4 categories

A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.

Tool callingVisionStreamingJSON mode
Context window262.1K tok
Max output209.7K tok
ReleasedDec 2, 2025
Knowledge cutoff
Text + ImageText
Input$0.15 / 1M tok
Output$0.15 / 1M tok
Cache read$0.015 / 1M tok
Intelligence5.5
Coding9.7
Agentic0.6

Design Arena — ranked #96 across 4 categories

A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.

Tool callingVisionStreamingJSON mode
Context window262.1K tok
Max output209.7K tok
ReleasedDec 2, 2025
Knowledge cutoff
Text + ImageText
Input$0.075 / 1M tok
Output$0.075 / 1M tok
Cache read$0.0075 / 1M tok
Intelligence5.5
Coding9.7
Agentic0.6

Design Arena — ranked #96 across 4 categories

Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.

Tool callingVisionStreamingJSON mode
Context window262.1K tok
Max output209.7K tok
ReleasedNov 1, 2024
Knowledge cutoffNov 1, 2024
Text + Image + FileText
Input$0.5 / 1M tok
Output$1.5 / 1M tok
Cache read$0.05 / 1M tok
Intelligence9.7
Coding20.1
Agentic2.4

Design Arena — ranked #61 across 8 categories

Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.

Tool callingVisionStreamingJSON mode
Context window262.1K tok
Max output209.7K tok
ReleasedDec 1, 2025
Knowledge cutoff
Text + Image + FileText
Input$0.25 / 1M tok
Output$0.75 / 1M tok
Cache read$0.025 / 1M tok
Intelligence9.7
Coding20.1
Agentic2.4

Design Arena — ranked #61 across 8 categories

Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...

Tool callingVisionStreamingJSON mode
Context window131.1K tok
Max output104.9K tok
ReleasedMay 7, 2025
Knowledge cutoffMar 31, 2025
Text + Image + FileText
Input$0.4 / 1M tok
Output$2 / 1M tok
Cache read$0.04 / 1M tok

Design Arena — ranked #80 across 6 categories

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...

Tool callingVisionStreamingJSON mode
Context window131.1K tok
Max output104.9K tok
ReleasedAug 13, 2025
Knowledge cutoffJun 30, 2025
Text + Image + FileText
Input$0.4 / 1M tok
Output$2 / 1M tok
Cache read$0.04 / 1M tok
Coding20.5
Agentic3.1

Design Arena — ranked #65 across 8 categories

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...

Tool callingVisionStreamingJSON mode
Context window131.1K tok
Max output104.9K tok
ReleasedAug 13, 2025
Knowledge cutoffJun 30, 2025
Text + Image + FileText
Input$0.2 / 1M tok
Output$1 / 1M tok
Cache read$0.02 / 1M tok
Coding20.5
Agentic3.1

Design Arena — ranked #65 across 8 categories

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

ReasoningTool callingVisionStreamingJSON mode
Context window262.1K tok
Max output209.7K tok
ReleasedApr 30, 2026
Knowledge cutoff
Text + Image + FileText
Input$1.5 / 1M tok
Output$7.5 / 1M tok
Effort: high, none
Intelligence14.9
Coding46.9
Agentic9.4

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

ReasoningTool callingVisionStreamingJSON mode
Context window262.1K tok
Max output209.7K tok
ReleasedApr 30, 2026
Knowledge cutoff
Text + Image + FileText
Input$0.75 / 1M tok
Output$3.75 / 1M tok
Effort: high, none
Intelligence14.9
Coding46.9
Agentic9.4

A 12B parameter model with a 128k token context length built by Mistral in collaboration with NVIDIA. The model is multilingual, supporting English, French, German, Spanish, Italian, Portuguese, Chinese, Japanese,...

Tool callingStreamingJSON mode
Context window131.1K tok
Max output16.4K tok
ReleasedJul 1, 2024
Knowledge cutoffApr 30, 2024
TextText
Input$0.019 / 1M tok
Output$0.03 / 1M tok

Mistral Small 3 is a 24B-parameter language model optimized for low-latency performance across common AI tasks. Released under the Apache 2.0 license, it features both pre-trained and instruction-tuned versions designed...

StreamingJSON mode
Context window32.8K tok
Max output16.4K tok
ReleasedJan 30, 2025
Knowledge cutoffOct 31, 2023
TextText
Input$0.05 / 1M tok
Output$0.08 / 1M tok

Mistral Small 3.1 24B Instruct is an upgraded variant of Mistral Small 3 (2501), featuring 24 billion parameters with advanced multimodal capabilities. It provides state-of-the-art performance in text-based reasoning and...

VisionStreaming
Context window128K tok
Max output102.4K tok
ReleasedMar 17, 2025
Knowledge cutoffOct 31, 2023
Text + ImageText
Input$0.351 / 1M tok
Output$0.555 / 1M tok

Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved function calling. Compared to the 3.1 release, version 3.2 significantly improves accuracy on...

Tool callingVisionStreamingJSON mode
Context window256K tok
Max output16.4K tok
ReleasedJun 20, 2025
Knowledge cutoffOct 31, 2023
Image + TextText
Input$0.075 / 1M tok
Output$0.2 / 1M tok

Design Arena — ranked #113 across 5 categories

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...

ReasoningTool callingVisionStreamingJSON mode
Context window262.1K tok
Max output209.7K tok
ReleasedMar 16, 2026
Knowledge cutoffJun 1, 2025
Text + ImageText
Input$0.15 / 1M tok
Output$0.6 / 1M tok
Cache read$0.015 / 1M tok
Effort: high, none
Intelligence11.5
Coding26.6
Agentic1.4

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...

ReasoningTool callingVisionStreamingJSON mode
Context window262.1K tok
Max output209.7K tok
ReleasedMar 16, 2026
Knowledge cutoff
Text + ImageText
Input$0.075 / 1M tok
Output$0.3 / 1M tok
Cache read$0.0075 / 1M tok
Effort: high, none
Intelligence11.5
Coding26.6
Agentic1.4

Mistral's official instruct fine-tuned version of [Mixtral 8x22B](/models/mistralai/mixtral-8x22b). It uses 39B active parameters out of 141B, offering unparalleled cost efficiency for its size. Its strengths include: - strong math, coding,...

Tool callingStreamingJSON mode
Context window65.5K tok
Max output52.4K tok
ReleasedApr 17, 2024
Knowledge cutoffJan 31, 2024
Text + FileText
Input$2 / 1M tok
Output$6 / 1M tok
Cache read$0.2 / 1M tok

Mistral Saba is a 24B-parameter language model specifically designed for the Middle East and South Asia, delivering accurate and contextually relevant responses while maintaining efficient performance. Trained on curated regional...

Tool callingStreamingJSON mode
Context window32.8K tok
Max output26.2K tok
ReleasedFeb 17, 2025
Knowledge cutoffSep 30, 2024
Text + FileText
Input$0.2 / 1M tok
Output$0.6 / 1M tok
Cache read$0.02 / 1M tok

Every Mistral AI model above, one subscription away.

UltraGPT puts this whole directory behind a single login and one flat monthly price — no separate bill per provider.