M

Mistral: mistral-nemo:free

mistral-nemo:free
128K context4.1K outToolsStructured
Released Jul 19, 2024Knowledge cutoff Apr 2024Updated Jun 18, 2026

A 12B parameter model with a 128k token context length built by Mistral in collaboration with NVIDIA. The model is multilingual, supporting English, French, German, Spanish, Italian, Portuguese, Chinese, Japanese, Korean, Arabic, and Hindi. It supports function calling and is released under the Apache 2.0 license.

Mode chatTokenizer MistralQuantization fp8mistralai/Mistral-Nemo-Instruct-2407

Pricing

Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Context window 128K tokensCompatible endpoints openaiVendor Mistral

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Supported parameters

All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.

ParameterProvidersDefault
frequency_penaltyAll providers-
logit_biasAll providers-
logprobsAll providers-
max_tokensAll providers-
min_pSome providers-
presence_penaltyAll providers-
repetition_penaltyAll providers-
response_formatAll providers-
seedAll providers-
stopAll providers-
structured_outputsAll providers-
temperatureAll providers0.3
tool_choiceSome providers-
toolsSome providers-
top_kAll providers-
top_logprobsAll providers-
top_pAll providers-

Frequently asked questions

Similar models