M

Meta: llama-3.3-70b:free

llama-3.3-70b:free
128K context128K outToolsStructured
Released Dec 6, 2024Knowledge cutoff Dec 2023Updated Jun 20, 2026

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model is optimized for multilingual dialogue use cases and outperforms many of the available open source and closed chat models on common industry benchmarks. Supported languages: English, German, French, Italian, Portuguese, Hindi, Spanish, and Thai. Model Card

Mode chatTokenizer Llama3Quantization fp8meta-llama/Llama-3.3-70B-Instruct

Pricing

Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Context window 128K tokensCompatible endpoints openaiVendor Meta

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Supported parameters

All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.

ParameterProvidersDefault
frequency_penaltyAll providers-
logit_biasSome providers-
logprobsSome providers-
max_tokensAll providers-
min_pSome providers-
presence_penaltyAll providers-
repetition_penaltyAll providers-
response_formatAll providers-
seedAll providers-
stopAll providers-
structured_outputsAll providers-
temperatureAll providers-
tool_choiceAll providers-
toolsAll providers-
top_kAll providers-
top_logprobsSome providers-
top_pAll providers-

Frequently asked questions

Similar models