Provider

IInclusionai

Every Inclusionai model UltraGPT can reach, kept in sync with the live catalog — pricing, context window and capabilities for each one.

6 models from Inclusionai Up to 262.1K tokens of context 3 free to run
Open the appBrowse every provider

6

Models tracked

1

Providers

6

Reasoning models

2

See images (vision)

262K

Largest context window

3

Free to run

Directory

All Inclusionai models.

Search and filter this provider's lineup — pricing per million tokens, context window, modalities, reasoning effort and independent benchmark scores where available.

6 of 6 models

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

ReasoningTool callingStreamingJSON mode
Context window262.1K tok
Max output32.8K tok
ReleasedJul 23, 2026
Knowledge cutoff
TextText
Input$0.021 / 1M tok
Output$0.063 / 1M tok
Cache read$0.0042 / 1M tok
Coding50.6
Agentic21.0

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...

ReasoningTool callingStreamingJSON mode
Context window262.1K tok
Max output235.9K tok
ReleasedAug 27, 2026
Knowledge cutoff
TextText
Input$0.06 / 1M tok
Output$0.18 / 1M tok
Cache read$0.012 / 1M tok

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...

ReasoningTool callingStreaming
Context window262.1K tok
Max output32.8K tok
ReleasedAug 27, 2026
Knowledge cutoff
TextText
InputFree
OutputFree

Ling 3.0 Flash Sante is a health and medicine-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for...

ReasoningTool callingStreaming
Context window262.1K tok
Max output32.8K tok
ReleasedSep 4, 2026
Knowledge cutoff
TextText
InputFree
OutputFree

Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...

ReasoningTool callingVisionStreamingJSON mode
Context window131.1K tok
Max output32.8K tok
ReleasedSep 10, 2026
Knowledge cutoff
Text + Image + VideoText
Input$0.06 / 1M tok
Output$0.18 / 1M tok
Cache read$0.012 / 1M tok
Intelligence24.8
Coding57.0
Agentic30.0

Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...

ReasoningTool callingVisionStreaming
Context window262.1K tok
Max output32.8K tok
ReleasedSep 10, 2026
Knowledge cutoff
Text + Image + VideoText
InputFree
OutputFree
Intelligence24.8
Coding57.0
Agentic30.0

Every Inclusionai model above, one subscription away.

UltraGPT puts this whole directory behind a single login and one flat monthly price — no separate bill per provider.