I
InclusionAI: ling-3.0-flash-fin:free
ling-3.0-flash-fin:free
262.1K context32.8K outReasoningToolsCacheStructured
Released Jul 23, 2026Updated Aug 31, 2026
Ling-3.0-flash is a 124B-parameter Mixture-of-Experts (MoE) model, with approximately 5.1B parameters activated per token. The model is designed with token efficiency and production-scale agentic inference as key priorities, enabling developers to complete more useful work within constrained token, latency, and serving-cost budgets.
Pricing
Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Context window 262.1K tokensCompatible endpoints openaiVendor InclusionAI
Uptime
Performance
Loading performance data...
Usage & Ranking
Loading usage...
Supported parameters
All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.
| Parameter | Providers | Default |
|---|---|---|
| frequency_penalty | All providers | - |
| include_reasoning | All providers | - |
| logit_bias | All providers | - |
| logprobs | All providers | - |
| max_tokens | All providers | - |
| min_p | All providers | - |
| presence_penalty | All providers | - |
| reasoning | All providers | - |
| repetition_penalty | All providers | - |
| response_format | All providers | - |
| seed | All providers | - |
| stop | All providers | - |
| temperature | All providers | - |
| tool_choice | All providers | - |
| tools | All providers | - |
| top_k | All providers | - |
| top_logprobs | All providers | - |
| top_p | All providers | - |