Every model, one list.

The full catalog UltraGPT draws from — context window, live pricing, capabilities and benchmarks for each model, kept in sync automatically. No guessing which model does what before you switch.

447 models across 52 providers
Open the appBrowse the directory

447

Models tracked

52

Providers

316

Reasoning models

277

See images (vision)

2,000K

Largest context window

22

Free to run

Directory

Search, filter, compare.

Every field the catalog tracks — pricing per million tokens, context window, input/output modalities, reasoning effort and independent benchmark scores where available.

447 of 447 models

Aion-2.0 is a variant of DeepSeek V3.2 optimized for immersive roleplaying and storytelling. It is particularly strong at introducing tension, crises, and conflict into stories, making narratives feel more engaging....

ReasoningTool callingStreamingJSON mode
Context window131.1K tok
Max output32.8K tok
ReleasedFeb 23, 2026
Knowledge cutoff
TextText
Input$0.8 / 1M tok
Output$1.6 / 1M tok
Cache read$0.2 / 1M tok

Aion-3.0 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses a collaborative generation process in which multiple specialized models each contribute...

ReasoningTool callingStreamingJSON mode
Context window131.1K tok
Max output32.8K tok
ReleasedJul 7, 2026
Knowledge cutoff
TextText
Input$3 / 1M tok
Output$6 / 1M tok
Cache read$0.75 / 1M tok

Aion-3.0 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the DeepSeek family of models. It uses a collaborative generation process in which multiple specialized models each...

ReasoningTool callingStreamingJSON mode
Context window131.1K tok
Max output32.8K tok
ReleasedJul 7, 2026
Knowledge cutoff
TextText
Input$0.7 / 1M tok
Output$1.4 / 1M tok
Cache read$0.18 / 1M tok

Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model...

Streaming
Context window32.8K tok
Max output29.5K tok
ReleasedFeb 4, 2025
Knowledge cutoffDec 31, 2023
TextText
Input$0.8 / 1M tok
Output$1.6 / 1M tok

Nova 2 Lite is a fast, cost-effective reasoning model for everyday workloads that can process text, images, and videos to generate text. Nova 2 Lite demonstrates standout capabilities in processing...

ReasoningTool callingVisionStreaming
Context window1M tok
Max output65.5K tok
ReleasedDec 2, 2025
Knowledge cutoff
Text + Image + Video + FileText
Input$0.3 / 1M tok
Output$2.5 / 1M tok
Coding23.0

Amazon Nova Lite 1.0 is a very low-cost multimodal model from Amazon that focused on fast processing of image, video, and text inputs to generate text output. Amazon Nova Lite...

Tool callingVisionStreaming
Context window300K tok
Max output5.1K tok
ReleasedDec 5, 2024
Knowledge cutoffOct 31, 2024
Text + ImageText
Input$0.06 / 1M tok
Output$0.24 / 1M tok

Amazon Nova Micro 1.0 is a text-only model that delivers the lowest latency responses in the Amazon Nova family of models at a very low cost. With a context length...

Tool callingStreaming
Context window128K tok
Max output5.1K tok
ReleasedDec 5, 2024
Knowledge cutoffOct 31, 2024
TextText
Input$0.035 / 1M tok
Output$0.14 / 1M tok

Amazon Nova Premier is the most capable of Amazon’s multimodal models for complex reasoning tasks and for use as the best teacher for distilling custom models.

Tool callingVisionStreaming
Context window1M tok
Max output32K tok
ReleasedOct 31, 2025
Knowledge cutoff
Text + ImageText
Input$2.5 / 1M tok
Output$12.5 / 1M tok
Cache read$0.625 / 1M tok

Design Arena — ranked #127 across 1 categories

Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide range of tasks. As of December...

Tool callingVisionStreaming
Context window300K tok
Max output5.1K tok
ReleasedDec 5, 2024
Knowledge cutoffOct 31, 2024
Text + ImageText
Input$0.8 / 1M tok
Output$3.2 / 1M tok

Design Arena — ranked #130 across 1 categories

Claude 3 Haiku is Anthropic's fastest and most compact model for near-instant responsiveness. Quick and accurate targeted performance. See the launch announcement and benchmark results [here](https://www.anthropic.com/news/claude-3-haiku) #multimodal

Tool callingVisionStreaming
Context window200K tok
Max output4.1K tok
ReleasedMar 13, 2024
Knowledge cutoffAug 31, 2023
Text + ImageText
Input$0.25 / 1M tok
Output$1.25 / 1M tok
Cache read$0.03 / 1M tok
Cache write$0.3 / 1M tok

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

ReasoningTool callingVisionStreamingJSON mode
Context window1M tok
Max output128K tok
ReleasedJun 7, 2026
Knowledge cutoff
Text + Image + FileText
Input$10 / 1M tok
Output$50 / 1M tok
Cache read$1 / 1M tok
Cache write$12.5 / 1M tok
Effort: max, xhigh, high, medium, low · always on
Intelligence49.7
Coding76.5
Agentic51.0

Design Arena — ranked #1 across 18 categories

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

ReasoningTool callingVisionStreamingJSON mode
Context window1M tok
Max output128K tok
ReleasedJun 9, 2026
Knowledge cutoff
Text + Image + FileText
Input$5 / 1M tok
Output$25 / 1M tok
Cache read$0.5 / 1M tok
Cache write$6.25 / 1M tok
Effort: max, xhigh, high, medium, low · always on
Intelligence49.7
Coding76.5
Agentic51.0

Design Arena — ranked #1 across 18 categories

Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...

ReasoningTool callingVisionStreamingJSON mode
Context window1M tok
Max output128K tok
ReleasedSep 1, 2026
Knowledge cutoffJun 1, 2026
Text + Image + FileText
Input$10 / 1M tok
Output$50 / 1M tok
Cache read$0.25 / 1M tok
Cache write$12.5 / 1M tok
Effort: max, xhigh, high, medium, low · always on
Intelligence53.4
Coding81.6
Agentic58.0

Design Arena — ranked #1 across 10 categories

Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...

ReasoningTool callingVisionStreamingJSON mode
Context window1M tok
Max output128K tok
ReleasedSep 1, 2026
Knowledge cutoff
Text + Image + FileText
Input$5 / 1M tok
Output$25 / 1M tok
Cache read$0.125 / 1M tok
Cache write$6.25 / 1M tok
Effort: max, xhigh, high, medium, low · always on
Intelligence53.4
Coding81.6
Agentic58.0

Design Arena — ranked #1 across 10 categories

This model always redirects to the latest model in the Claude Fable family.

ReasoningTool callingVisionStreamingJSON mode
Context window1M tok
Max output128K tok
ReleasedJun 9, 2026
Knowledge cutoff
Text + Image + FileText
Input$10 / 1M tok
Output$50 / 1M tok
Cache read$0.25 / 1M tok
Cache write$12.5 / 1M tok
Effort: max, xhigh, high, medium, low · always on

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...

ReasoningTool callingVisionStreamingJSON mode
Context window200K tok
Max output64K tok
ReleasedOct 15, 2025
Knowledge cutoffFeb 28, 2025
Text + Image + FileText
Input$1 / 1M tok
Output$5 / 1M tok
Cache read$0.1 / 1M tok
Cache write$1.25 / 1M tok
Intelligence17.6
Coding43.9
Agentic10.3

Design Arena — ranked #41 across 8 categories

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...

ReasoningTool callingVisionStreamingJSON mode
Context window200K tok
Max output64K tok
ReleasedOct 15, 2025
Knowledge cutoff
Text + Image + FileText
Input$0.5 / 1M tok
Output$2.5 / 1M tok
Cache read$0.05 / 1M tok
Cache write$0.625 / 1M tok
Intelligence17.6
Coding43.9
Agentic10.3

Design Arena — ranked #41 across 8 categories

This model always redirects to the latest model in the Claude Haiku family.

ReasoningTool callingVisionStreamingJSON mode
Context window200K tok
Max output64K tok
ReleasedApr 27, 2026
Knowledge cutoff
Text + Image + FileText
Input$1 / 1M tok
Output$5 / 1M tok
Cache read$0.1 / 1M tok
Cache write$1.25 / 1M tok

Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...

ReasoningTool callingVisionStreaming
Context window200K tok
Max output32K tok
ReleasedMay 22, 2025
Knowledge cutoffJan 31, 2025
Image + Text + FileText
Input$15 / 1M tok
Output$75 / 1M tok
Cache read$1.5 / 1M tok
Cache write$18.75 / 1M tok

Design Arena — ranked #49 across 7 categories

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

ReasoningTool callingVisionStreaming
Context window200K tok
Max output32K tok
ReleasedAug 5, 2025
Knowledge cutoffJan 31, 2025
Image + Text + FileText
Input$15 / 1M tok
Output$75 / 1M tok
Cache read$1.5 / 1M tok
Cache write$18.75 / 1M tok

Design Arena — ranked #26 across 8 categories

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

ReasoningTool callingVisionStreamingJSON mode
Context window200K tok
Max output32K tok
ReleasedAug 5, 2025
Knowledge cutoffJan 31, 2025
Image + Text + FileText
Input$7.5 / 1M tok
Output$37.5 / 1M tok
Cache read$0.75 / 1M tok
Cache write$9.38 / 1M tok

Design Arena — ranked #26 across 8 categories

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...

ReasoningTool callingVisionStreamingJSON mode
Context window200K tok
Max output64K tok
ReleasedNov 24, 2025
Knowledge cutoffMay 1, 2025
File + Image + TextText
Input$5 / 1M tok
Output$25 / 1M tok
Cache read$0.5 / 1M tok
Cache write$6.25 / 1M tok

Design Arena — ranked #15 across 12 categories

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...

ReasoningTool callingVisionStreamingJSON mode
Context window200K tok
Max output64K tok
ReleasedNov 24, 2025
Knowledge cutoff
File + Image + TextText
Input$2.5 / 1M tok
Output$12.5 / 1M tok
Cache read$0.25 / 1M tok
Cache write$3.13 / 1M tok

Design Arena — ranked #15 across 12 categories

Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...

ReasoningTool callingVisionStreamingJSON mode
Context window1M tok
Max output128K tok
ReleasedFeb 4, 2026
Knowledge cutoffMay 31, 2025
Text + Image + FileText
Input$5 / 1M tok
Output$25 / 1M tok
Cache read$0.5 / 1M tok
Cache write$6.25 / 1M tok
Effort: max, high, medium, low

Design Arena — ranked #8 across 13 categories

Every model above, one subscription away.

UltraGPT puts this whole directory behind a single login and one flat monthly price — no separate bill per provider.