Provider

Inception

Every Inception model UltraGPT can reach, kept in sync with the live catalog — pricing, context window and capabilities for each one.

2 models from Inception Up to 260K tokens of context
Open the appBrowse every provider

2

Models tracked

1

Providers

2

Reasoning models

0

See images (vision)

260K

Largest context window

0

Free to run

Directory

All Inception models.

Search and filter this provider's lineup — pricing per million tokens, context window, modalities, reasoning effort and independent benchmark scores where available.

2 of 2 models

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...

ReasoningTool callingStreamingJSON mode
Context window128K tok
Max output50K tok
ReleasedFeb 24, 2026
Knowledge cutoffJan 1, 2025
TextText
Input$0.25 / 1M tok
Output$0.75 / 1M tok
Cache read$0.025 / 1M tok
Effort: high, medium, low, none
Intelligence11.5
Coding31.1
Agentic4.0

Design Arena — ranked #64 across 8 categories

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

ReasoningTool callingStreamingJSON mode
Context window260K tok
Max output65.5K tok
ReleasedSep 8, 2026
Knowledge cutoffNov 1, 2025
TextText
Input$0.04 / 1M tok
Output$0.15 / 1M tok
Cache read$0.004 / 1M tok
Effort: high, medium, low, none

Every Inception model above, one subscription away.

UltraGPT puts this whole directory behind a single login and one flat monthly price — no separate bill per provider.