Provider

MiniMax

Every MiniMax model UltraGPT can reach, kept in sync with the live catalog — pricing, context window and capabilities for each one.

9 models from MiniMax Up to 1.05M tokens of context
Open the appBrowse every provider

9

Models tracked

1

Providers

7

Reasoning models

3

See images (vision)

1,049K

Largest context window

0

Free to run

Directory

All MiniMax models.

Search and filter this provider's lineup — pricing per million tokens, context window, modalities, reasoning effort and independent benchmark scores where available.

9 of 9 models

MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference. It leverages a hybrid Mixture-of-Experts (MoE) architecture paired with a custom "lightning attention" mechanism, allowing it...

ReasoningTool callingStreaming
Context window1M tok
Max output40K tok
ReleasedJun 17, 2025
Knowledge cutoffJun 30, 2024
TextText
Input$0.55 / 1M tok
Output$2.2 / 1M tok

MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows. With 10 billion activated parameters (230 billion total), it delivers near-frontier intelligence across general reasoning,...

ReasoningTool callingStreamingJSON mode
Context window204.8K tok
Max output131.1K tok
ReleasedOct 27, 2025
Knowledge cutoff
TextText
Input$0.255 / 1M tok
Output$1.02 / 1M tok

Design Arena — ranked #55 across 7 categories

MiniMax M2-her is a dialogue-first large language model built for immersive roleplay, character-driven chat, and expressive multi-turn conversations. Designed to stay consistent in tone and personality, it supports rich message...

Streaming
Context window65.5K tok
Max output2.0K tok
ReleasedJan 23, 2026
Knowledge cutoff
TextText
Input$0.3 / 1M tok
Output$1.2 / 1M tok
Cache read$0.03 / 1M tok

MiniMax-M2.1 is a lightweight, state-of-the-art large language model optimized for coding, agentic workflows, and modern application development. With only 10 billion activated parameters, it delivers a major jump in real-world...

ReasoningTool callingStreamingJSON mode
Context window204.8K tok
Max output131.1K tok
ReleasedDec 23, 2025
Knowledge cutoff
TextText
Input$0.3 / 1M tok
Output$1.2 / 1M tok
Cache read$0.03 / 1M tok
Cache write$0.375 / 1M tok

Design Arena — ranked #41 across 7 categories

MiniMax-M2.5 is a SOTA large language model designed for real-world productivity. Trained in a diverse range of complex real-world digital working environments, M2.5 builds upon the coding expertise of M2.1...

ReasoningTool callingStreamingJSON mode
Context window204.8K tok
Max output128K tok
ReleasedFeb 12, 2026
Knowledge cutoff
TextText
Input$0.27 / 1M tok
Output$1.08 / 1M tok
Cache read$0.027 / 1M tok
Cache write$0.375 / 1M tok

Design Arena — ranked #40 across 7 categories

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...

ReasoningTool callingStreamingJSON mode
Context window204.8K tok
Max output131.1K tok
ReleasedMar 18, 2026
Knowledge cutoff
TextText
Input$0.3 / 1M tok
Output$1.2 / 1M tok
Cache read$0.06 / 1M tok
Cache write$0.375 / 1M tok
Intelligence23.2
Coding52.6
Agentic16.8

Design Arena — ranked #34 across 8 categories

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

ReasoningTool callingVisionStreamingJSON mode
Context window1.05M tok
Max output512K tok
ReleasedJun 1, 2026
Knowledge cutoff
Text + Image + VideoText
Input$0.3 / 1M tok
Output$1.2 / 1M tok
Cache read$0.06 / 1M tok
Intelligence29.6
Coding58.6
Agentic30.8

Design Arena — ranked #11 across 15 categories

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

ReasoningTool callingVisionStreamingJSON mode
Context window524.3K tok
Max output471.9K tok
ReleasedMay 31, 2026
Knowledge cutoff
Text + Image + VideoText
Input$0.3 / 1M tok
Output$1.2 / 1M tok
Cache read$0.06 / 1M tok
Intelligence29.6
Coding58.6
Agentic30.8

Design Arena — ranked #11 across 15 categories

MiniMax-01 is a combines MiniMax-Text-01 for text generation and MiniMax-VL-01 for image understanding. It has 456 billion parameters, with 45.9 billion parameters activated per inference, and can handle a context...

VisionStreaming
Context window1.00M tok
Max output900.2K tok
ReleasedJan 15, 2025
Knowledge cutoffMar 31, 2024
Text + ImageText
Input$0.2 / 1M tok
Output$1.1 / 1M tok

Every MiniMax model above, one subscription away.

UltraGPT puts this whole directory behind a single login and one flat monthly price — no separate bill per provider.