Meta: Llama 4 Scout
meta-llama/llama-4-scout
Description
Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input (text and image) and multilingual output (text and code) across 12 supported languages. Designed for assistant-style interaction and visual reasoning, Scout uses 16 experts per forward pass and features a context length of 10 million tokens, with a training corpus of ~40 trillion tokens.
Built for high efficiency and local or commercial deployment, Llama 4 Scout incorporates early fusion for seamless modality integration. It is instruction-tuned for use in multilingual chat, captioning, and image understanding tasks. Released under the Llama 4 Community License, it was last trained on data up to August 2024 and launched publicly on April 5, 2025.
How this model compares
Overall covers the full catalog. By plan covers only models available on that tier (same rules as available models in your list). Position on min–average–max. Prices use the higher of prompt or completion per token, shown per 1M tokens.
Price (per 1M tokens)
Min
Max
This model
339 models in this groupPrice (per 1M tokens)
- Min
- $0.04
- Avg
- $12.395447
- Max
- $750.00
This model: $0.30 / 1M tokens
Context length (tokens)
Min
Max
This model
339 models in this groupContext length (tokens)
- Min
- 4,095 tokens
- Avg
- 379,884.782 tokens
- Max
- 10,000,000 tokens
This model: 10,000,000 tokens
Capabilities
Text + Image → TextContext: 327,680 tokens
Input:
TextImage
Output:
Text