A

Alibaba: qwen3-next-80b-a3b-instruct:free

qwen3-next-80b-a3b-instruct:free
13.1万 コンテキスト6.6万 出力ツールキャッシュ構造化
リリース日 Sep 11, 2025知識のカットオフ Sep 2025更新日 Jun 18, 2026

Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual use, while remaining robust on alignment and formatting. Compared with prior Qwen3 instruct variants, it focuses on higher throughput and stability on ultra-long inputs and multi-turn dialogues, making it well-suited for RAG, tool use, and agentic workflows that require consistent final answers rather than visible chain-of-thought. The model employs scaling-efficient training and decoding to improve parameter efficiency and inference speed, and has been validated on a broad set of public benchmarks where it reaches or approaches larger Qwen3 systems in several categories while outperforming earlier mid-sized baselines. It is best used as a general assistant, code helper, and long-context task solver in production settings where deterministic, instruction-following outputs are preferred.

モード chatトークナイザー Qwen3量子化 fp8Qwen/Qwen3-Next-80B-A3B-Instruct

料金

入力料金
$0.00/ 100万トークン
出力料金
$0.00/ 100万トークン
コンテキストウィンドウ 131.1K トークン対応エンドポイント openaiベンダー Alibaba

稼働率

パフォーマンス

パフォーマンスデータを読み込み中...

使用状況とランキング

使用状況を読み込み中...

対応パラメータ

すべてのプロバイダー = このモデルを提供するすべてのアップストリームが対応しています。一部のプロバイダー = リクエストを処理するアップストリームによって異なります。デフォルト = 未設定時に送信される値です。

パラメータプロバイダーデフォルト
frequency_penaltyすべてのプロバイダー-
logit_bias一部のプロバイダー-
logprobsすべてのプロバイダー-
max_tokensすべてのプロバイダー-
min_p一部のプロバイダー-
presence_penaltyすべてのプロバイダー-
repetition_penaltyすべてのプロバイダー-
response_formatすべてのプロバイダー-
seedすべてのプロバイダー-
stopすべてのプロバイダー-
structured_outputsすべてのプロバイダー-
temperatureすべてのプロバイダー-
tool_choiceすべてのプロバイダー-
toolsすべてのプロバイダー-
top_kすべてのプロバイダー-
top_logprobsすべてのプロバイダー-
top_pすべてのプロバイダー-

よくあるご質問

類似モデル