Meta: Llama 4 Maverick
Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...
Intelligence overview
1.05M tok
Context window
16.4K tok
Max output
What Meta: Llama 4 Maverick can do.
Multimodal
Image understanding
Reads and reasons over images alongside text in the same prompt.
Development
Tool calling
Calls external functions mid-response, then continues reasoning from the result.
Structured output
Returns responses constrained to a JSON schema, ready to parse without cleanup.
Streaming
Streams tokens as they're generated instead of waiting on the full response.
Performance
Design Arena — ranked #110 across 6 categories
Pay for what you use.
Every model in UltraGPT runs behind one flat subscription — this is the underlying per-token cost the catalog tracks.
$0.188 / 1M tok
$0.652 / 1M tok
Built for serious work.
Build
Generate and refactor code, with the ability to run and verify tool calls and return structured, parseable output.
Research
Holds up to 1.05M tokens of source material in context without losing the thread.
Create
Reads images alongside a prompt to ground creative and design work in what's actually on screen.
Ask Meta: Llama 4 Maverick anything.
A preview of the real UltraGPT composer — the model you see here is the model you get.
Continues in the UltraGPT app — one login, every model.
How Meta: Llama 4 Maverick stacks up.
The closest models in the catalog by provider, capability and score — real entries, pulled from the same directory.
Technical details.
Specification
- Context window
- 1.05M tokens
- Max output
- 16.4K tokens
- Input modalities
- Text + Image
- Output modalities
- Text
- Knowledge cutoff
- Aug 31, 2024
- Release date
- Apr 5, 2025
- Streaming
- Supported
- Tool calling
- Supported
- JSON mode
- Supported
More from Meta, and everything else.
This model is one entry in the full UltraGPT catalog — browse the rest, compare pricing and context windows, or filter by capability.
