Provider

Meta

Every Meta model UltraGPT can reach, kept in sync with the live catalog — pricing, context window and capabilities for each one.

8 models from Meta Up to 1.31M tokens of context
Open the appBrowse every provider

8

Models tracked

1

Providers

0

Reasoning models

3

See images (vision)

1,311K

Largest context window

0

Free to run

Directory

All Meta models.

Search and filter this provider's lineup — pricing per million tokens, context window, modalities, reasoning effort and independent benchmark scores where available.

8 of 8 models

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 70B instruct-tuned version is optimized for high quality dialogue usecases. It has demonstrated strong...

Tool callingStreamingJSON mode
Context window131.1K tok
Max output16.4K tok
ReleasedJul 23, 2024
Knowledge cutoffDec 31, 2023
TextText
Input$0.4 / 1M tok
Output$0.4 / 1M tok

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient. It has demonstrated strong performance compared to...

Tool callingStreamingJSON mode
Context window131.1K tok
Max output118.0K tok
ReleasedJul 23, 2024
Knowledge cutoffDec 31, 2023
TextText
Input$0.05 / 1M tok
Output$0.08 / 1M tok
Cache read$0.025 / 1M tok
Coding5.4

Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows it to operate...

Streaming
Context window60K tok
Max output54K tok
ReleasedSep 25, 2024
Knowledge cutoffDec 31, 2023
TextText
Input$0.027 / 1M tok
Output$0.201 / 1M tok

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...

StreamingJSON mode
Context window131.1K tok
Max output118.0K tok
ReleasedSep 25, 2024
Knowledge cutoffDec 31, 2023
TextText
Input$0.05 / 1M tok
Output$0.33 / 1M tok

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...

Tool callingStreamingJSON mode
Context window131.1K tok
Max output16.4K tok
ReleasedDec 6, 2024
Knowledge cutoffDec 31, 2023
TextText
Input$0.1 / 1M tok
Output$0.32 / 1M tok
Coding11.9

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...

Tool callingVisionStreamingJSON mode
Context window1.05M tok
Max output16.4K tok
ReleasedApr 5, 2025
Knowledge cutoffAug 31, 2024
Text + ImageText
Input$0.188 / 1M tok
Output$0.652 / 1M tok
Intelligence9.3
Coding16.3
Agentic0.6

Design Arena — ranked #110 across 6 categories

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...

Tool callingVisionStreamingJSON mode
Context window1.31M tok
Max output16.4K tok
ReleasedApr 5, 2025
Knowledge cutoffAug 31, 2024
Text + ImageText
Input$0.1 / 1M tok
Output$0.3 / 1M tok
Intelligence6.5
Coding8.2
Agentic0.5

Design Arena — ranked #117 across 5 categories

Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...

VisionStreamingJSON mode
Context window163.8K tok
Max output16.4K tok
ReleasedApr 30, 2025
Knowledge cutoffAug 31, 2024
Image + TextText
Input$0.18 / 1M tok
Output$0.18 / 1M tok

Every Meta model above, one subscription away.

UltraGPT puts this whole directory behind a single login and one flat monthly price — no separate bill per provider.