inclusionAI: Ling 3.0 Flash Fin (free)
Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...
Intelligence overview
262.1K tok
Context window
32.8K tok
Max output
What inclusionAI: Ling 3.0 Flash Fin (free) can do.
Reasoning
Extended reasoning
Works through problems step by step before answering, trading latency for accuracy on harder tasks.
Development
Tool calling
Calls external functions mid-response, then continues reasoning from the result.
Streaming
Streams tokens as they're generated instead of waiting on the full response.
Pay for what you use.
Every model in UltraGPT runs behind one flat subscription — this is the underlying per-token cost the catalog tracks.
Free
Free
Built for serious work.
Build
Generate and refactor code, with the ability to run and verify tool calls.
Think
Works through multi-step problems with extended reasoning before it answers.
Research
Holds up to 262.1K tokens of source material in context without losing the thread.
Automate
Chains tool calls across multi-step workflows, reasoning between each one.
Ask inclusionAI: Ling 3.0 Flash Fin (free) anything.
A preview of the real UltraGPT composer — the model you see here is the model you get.
Continues in the UltraGPT app — one login, every model.
How inclusionAI: Ling 3.0 Flash Fin (free) stacks up.
The closest models in the catalog by provider, capability and score — real entries, pulled from the same directory.
Technical details.
Specification
- Context window
- 262.1K tokens
- Max output
- 32.8K tokens
- Input modalities
- Text
- Output modalities
- Text
- Release date
- Aug 27, 2026
- Streaming
- Supported
- Tool calling
- Supported
- JSON mode
- Not supported
More from Inclusionai, and everything else.
This model is one entry in the full UltraGPT catalog — browse the rest, compare pricing and context windows, or filter by capability.
