A

Alibaba: qwen3.5-flash:free

qwen3.5-flash:free
1M contesto65.5K outputRagionamentoStrumentiVisioneVideoStrutturato
Rilasciato Feb 25, 2026Aggiornato Aug 6, 2026

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the 3 series, these models deliver a leap forward in performance for both pure text and multimodal tasks, offering fast response times while balancing inference speed and overall performance.

Modalità chatTokenizer Qwen3

Prezzi

Prezzo di input
$0.00/ 1 M token
Prezzo di output
$0.00/ 1 M token
Endpoint compatibili openaiProvider Alibaba

Disponibilità

Performance

Caricamento dati di performance...

Utilizzo e classifica

Caricamento utilizzo...

Parametri supportati

Tutti i provider = supportato da ogni upstream che serve questo modello. Alcuni provider = dipende dall'upstream che gestisce la richiesta. Predefinito = il valore inviato quando non lo imposti.

ParametroProviderPredefinito
frequency_penaltyTutti i providerNon inviato per impostazione predefinita
include_reasoningTutti i provider-
max_tokensTutti i provider-
presence_penaltyTutti i provider-
reasoningTutti i provider-
response_formatTutti i provider-
seedTutti i provider-
stopTutti i provider-
structured_outputsTutti i provider-
temperatureTutti i providerNon inviato per impostazione predefinita
tool_choiceTutti i provider-
toolsTutti i provider-
top_kTutti i provider-
top_pTutti i providerNon inviato per impostazione predefinita

Domande frequenti

Modelli simili