A

Alibaba: qwen3.5-122b-a10b:free

qwen3.5-122b-a10b:free
262.1K context235.9K outReasoningToolsVisionVideoStructured
Released Feb 25, 2026Updated Jun 18, 2026

The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of overall performance, this model is second only to Qwen3.5-397B-A17B. Its text capabilities significantly outperform those of Qwen3-235B-2507, and its visual capabilities surpass those of Qwen3-VL-235B.

Mode chatTokenizer Qwen3Quantization fp8Qwen/Qwen3.5-122B-A10B

All providers for this model are busy right now

Every upstream provider has hit its rate limit. The model comes back automatically once limits lift, usually within hours. Try again in a little while or switch to another model.

Request this model on Discord

Pricing

Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Context window 262.1K tokensCompatible endpoints openaiVendor Alibaba

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Supported parameters

All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.

ParameterProvidersDefault
frequency_penaltyAll providersNot sent by default
include_reasoningAll providers-
logit_biasSome providers-
logprobsAll providers-
max_tokensAll providers-
min_pSome providers-
presence_penaltyAll providersNot sent by default
reasoningAll providers-
repetition_penaltyAll providersNot sent by default
response_formatAll providers-
seedAll providers-
stopAll providers-
structured_outputsAll providers-
temperatureAll providers0.6
tool_choiceAll providers-
toolsAll providers-
top_kAll providers20
top_logprobsAll providers-
top_pAll providers0.95

Frequently asked questions

Similar models