I

InclusionAI: ling-3.1-flash:free

ling-3.1-flash:free
262.1K context32.8K outReasoningTools
Released Oct 2, 2026Updated Oct 6, 2026

Ling 3.1 Flash is a hybrid reasoning mixture-of-experts model from inclusionAI, with 25B active parameters out of 560B total.

Mode chatTokenizer Other

All providers for this model are busy right now

Every upstream provider has hit its rate limit. The model comes back automatically once limits lift, usually within hours. Try again in a little while or switch to another model.

Request this model on Discord

Pricing

Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Context window 262.1K tokensCompatible endpoints openaiVendor InclusionAI

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Supported parameters

All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.

ParameterProvidersDefault
frequency_penaltySome providers-
include_reasoningSome providers-
logprobsSome providers-
max_tokensSome providers-
presence_penaltySome providers-
reasoningSome providers-
repetition_penaltySome providers-
seedSome providers-
stopSome providers-
temperatureSome providers-
tool_choiceSome providers-
toolsSome providers-
top_kSome providers-
top_logprobsSome providers-
top_pSome providers-

Frequently asked questions

Similar models