A

Alibaba: qwen3-32b:free

qwen3-32b:free
13.1万 コンテキスト12.8万 出力推論ツール構造化
リリース日 Apr 28, 2025知識のカットオフ Mar 2025更新日 Jun 18, 2026

A flagship dense model in the Qwen3 series, released by Alibaba Cloud on April 29, 2025. Featuring 32.8 billion parameters and a 64-layer Transformer architecture, it stands as a high-performance "mid-size" pillar of the Qwen family. It natively supports Dual-Mode (Thinking/Non-Thinking) switching, delivering logical reasoning depth comparable to previous-generation 72B or even 110B models within a 30B-scale footprint. With exceptional scores in AIME 2025 and LiveCodeBench, it is a premier choice for developers seeking a balance between inference speed and cognitive rigor.

モード chatトークナイザー Qwen3量子化 fp8Qwen/Qwen3-32B

このモデルのプロバイダーは現在すべて混雑しています

すべての上流プロバイダーがレート制限に達しました。制限が解除されるとモデルは自動的に復帰します。通常は数時間以内です。しばらくしてから再試行するか、別のモデルに切り替えてください。

このモデルをDiscordでリクエスト

料金

入力料金
$0.00/ 100万トークン
出力料金
$0.00/ 100万トークン
コンテキストウィンドウ 131.1K トークン対応エンドポイント openaiベンダー Alibaba

稼働率

パフォーマンス

パフォーマンスデータを読み込み中...

使用状況とランキング

使用状況を読み込み中...

対応パラメータ

すべてのプロバイダー = このモデルを提供するすべてのアップストリームが対応しています。一部のプロバイダー = リクエストを処理するアップストリームによって異なります。デフォルト = 未設定時に送信される値です。

パラメータプロバイダーデフォルト
frequency_penaltyすべてのプロバイダー-
include_reasoningすべてのプロバイダー-
logit_biasすべてのプロバイダー-
max_tokensすべてのプロバイダー-
min_pすべてのプロバイダー-
presence_penaltyすべてのプロバイダー-
reasoningすべてのプロバイダー-
repetition_penaltyすべてのプロバイダー-
response_formatすべてのプロバイダー-
seedすべてのプロバイダー-
stopすべてのプロバイダー-
structured_outputsすべてのプロバイダー-
temperatureすべてのプロバイダー-
tool_choiceすべてのプロバイダー-
toolsすべてのプロバイダー-
top_kすべてのプロバイダー-
top_pすべてのプロバイダー-

よくあるご質問

類似モデル