I
InclusionAI: ling-3.0-flash-fin:free
ling-3.0-flash-fin:free
262.1K contesto32.8K outputRagionamentoStrumentiCacheStrutturato
Rilasciato Jul 23, 2026Aggiornato Aug 31, 2026
Ling-3.0-flash is a 124B-parameter Mixture-of-Experts (MoE) model, with approximately 5.1B parameters activated per token. The model is designed with token efficiency and production-scale agentic inference as key priorities, enabling developers to complete more useful work within constrained token, latency, and serving-cost budgets.
Prezzi
Prezzo di input
$0.00/ 1 M token
Prezzo di output
$0.00/ 1 M token
Finestra di contesto 262.1K tokenEndpoint compatibili openaiProvider InclusionAI
Disponibilità
Performance
Caricamento dati di performance...
Utilizzo e classifica
Caricamento utilizzo...
Parametri supportati
Tutti i provider = supportato da ogni upstream che serve questo modello. Alcuni provider = dipende dall'upstream che gestisce la richiesta. Predefinito = il valore inviato quando non lo imposti.
| Parametro | Provider | Predefinito |
|---|---|---|
| frequency_penalty | Tutti i provider | - |
| include_reasoning | Tutti i provider | - |
| logit_bias | Tutti i provider | - |
| logprobs | Tutti i provider | - |
| max_tokens | Tutti i provider | - |
| min_p | Tutti i provider | - |
| presence_penalty | Tutti i provider | - |
| reasoning | Tutti i provider | - |
| repetition_penalty | Tutti i provider | - |
| response_format | Tutti i provider | - |
| seed | Tutti i provider | - |
| stop | Tutti i provider | - |
| temperature | Tutti i provider | - |
| tool_choice | Tutti i provider | - |
| tools | Tutti i provider | - |
| top_k | Tutti i provider | - |
| top_logprobs | Tutti i provider | - |
| top_p | Tutti i provider | - |