A

Alibaba: qwen3-14b:free

qwen3-14b:free
131.1K context41.0K outReasoningToolsStructured
Released Apr 28, 2025Knowledge cutoff Mar 2025Updated Jun 20, 2026

A mid-range dense model of the Qwen3 series released by Alibaba Cloud on April 29, 2025. With 14.8 billion parameters (approx. 13.2B non-embedding), it is the "sweet spot" model designed for mid-tier servers and high-performance workstations. It features the native Dual-Mode (Thinking/Non-Thinking), delivering logical reasoning that surpasses the previous Qwen2.5-32B and even some 70B models while maintaining the agility of a 14B scale. Trained on 36 trillion tokens, it excels in 119 languages and serves as a premier backbone for enterprise-grade AI agents.

Mode chatTokenizer Qwen3Quantization fp8Qwen/Qwen3-14B

All providers for this model are busy right now

Every upstream provider has hit its rate limit. The model comes back automatically once limits lift, usually within hours. Try again in a little while or switch to another model.

Request this model on Discord

Pricing

Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Context window 131.1K tokensCompatible endpoints openaiVendor Alibaba

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Supported parameters

All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.

ParameterProvidersDefault
frequency_penaltyAll providers-
include_reasoningAll providers-
logit_biasSome providers-
logprobsSome providers-
max_tokensAll providers-
min_pSome providers-
presence_penaltyAll providers-
reasoningAll providers-
repetition_penaltyAll providers-
response_formatAll providers-
seedAll providers-
stopAll providers-
structured_outputsAll providers-
temperatureAll providers-
tool_choiceAll providers-
toolsAll providers-
top_kAll providers-
top_logprobsSome providers-
top_pAll providers-

Frequently asked questions

Similar models