A

Alibaba: qwen3.5-35b-a3b:free

qwen3.5-35b-a3b:free
262.1K context65.5K outReasoningToolsParallel toolsVisionVideoCacheStructured
Released Feb 25, 2026Updated Jun 18, 2026

The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. Its overall performance is comparable to that of the Qwen3.5-27B.

Mode chatTokenizer Qwen3Qwen/Qwen3.5-35B-A3B

Pricing

Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Context window 262.1K tokensCompatible endpoints openaiVendor Alibaba

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Supported parameters

All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.

ParameterProvidersDefault
frequency_penaltyAll providersNot sent by default
include_reasoningAll providers-
logit_biasSome providers-
logprobsAll providers-
max_tokensAll providers-
min_pSome providers-
presence_penaltyAll providersNot sent by default
reasoningAll providers-
repetition_penaltyAll providersNot sent by default
response_formatAll providers-
seedAll providers-
stopAll providers-
structured_outputsAll providers-
temperatureAll providers1
tool_choiceAll providers-
toolsAll providers-
top_kAll providers20
top_logprobsAll providers-
top_pAll providers0.95

Frequently asked questions

Similar models