I
InclusionAI: ling-3.1-flash:free
ling-3.1-flash:free
262.1K context32.8K outReasoningTools
Released Oct 2, 2026Updated Oct 6, 2026
Ling 3.1 Flash is a hybrid reasoning mixture-of-experts model from inclusionAI, with 25B active parameters out of 560B total.
Mode chatTokenizer Other
All providers for this model are busy right now
Every upstream provider has hit its rate limit. The model comes back automatically once limits lift, usually within hours. Try again in a little while or switch to another model.
Request this model on DiscordPricing
Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Context window 262.1K tokensCompatible endpoints openaiVendor InclusionAI
Uptime
Performance
Loading performance data...
Usage & Ranking
Loading usage...
Supported parameters
All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.
| Parameter | Providers | Default |
|---|---|---|
| frequency_penalty | Some providers | - |
| include_reasoning | Some providers | - |
| logprobs | Some providers | - |
| max_tokens | Some providers | - |
| presence_penalty | Some providers | - |
| reasoning | Some providers | - |
| repetition_penalty | Some providers | - |
| seed | Some providers | - |
| stop | Some providers | - |
| temperature | Some providers | - |
| tool_choice | Some providers | - |
| tools | Some providers | - |
| top_k | Some providers | - |
| top_logprobs | Some providers | - |
| top_p | Some providers | - |