M

Meta: llama-4-scout:free

llama-4-scout:free
128K context8.2K outToolsVisionStructured
Released Apr 5, 2025Knowledge cutoff Aug 2024Updated Jun 20, 2026

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input (text and image) and multilingual output (text and code) across 12 supported languages. Designed for assistant-style interaction and visual reasoning, Scout uses 16 experts per forward pass and features a context length of 10 million tokens, with a training corpus of ~40 trillion tokens. Built for high efficiency and local or commercial deployment, Llama 4 Scout incorporates early fusion for seamless modality integration. It is instruction-tuned for use in multilingual chat, captioning, and image understanding tasks. Released under the Llama 4 Community License, it was last trained on data up to August 2024 and launched publicly on April 5, 2025.

Mode chatTokenizer Llama4Quantization bf16meta-llama/Llama-4-Scout-17B-16E-Instruct

Pricing

Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Context window 128K tokensCompatible endpoints openaiVendor Meta

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Supported parameters

All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.

ParameterProvidersDefault
frequency_penaltyAll providers-
logit_biasSome providers-
max_tokensAll providers-
min_pSome providers-
presence_penaltyAll providers-
repetition_penaltyAll providers-
response_formatAll providers-
seedAll providers-
stopAll providers-
structured_outputsAll providers-
temperatureAll providers-
tool_choiceAll providers-
toolsAll providers-
top_kAll providers-
top_pAll providers-

Frequently asked questions

Similar models