N

NVIDIA: nemotron-3-super-120b-a12b:free

nemotron-3-super-120b-a12b:free
25.6万 コンテキスト1.6万 出力推論ツール構造化
リリース日 Mar 11, 2026更新日 Jun 20, 2026

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer Mixture-of-Experts architecture with multi-token prediction (MTP), it delivers over 50% higher token generation compared to leading open models. The model features a 1M token context window for long-term agent coherence, cross-document reasoning, and multi-step task planning. Latent MoE enables calling 4 experts for the inference cost of only one, improving intelligence and generalization. Multi-environment RL training across 10+ environments delivers leading accuracy on benchmarks including AIME 2025, TerminalBench, and SWE-Bench Verified. Fully open with weights, datasets, and recipes under the NVIDIA Open License, Nemotron 3 Super allows easy customization and secure deployment anywhere — from workstation to cloud.

モード chatトークナイザー Othernvidia/NVIDIA-Nemotron-3-Super-120B-A12B-FP8

料金

入力料金
$0.00/ 100万トークン
出力料金
$0.00/ 100万トークン
コンテキストウィンドウ 256K トークン対応エンドポイント openaiベンダー NVIDIA

稼働率

パフォーマンス

パフォーマンスデータを読み込み中...

使用状況とランキング

使用状況を読み込み中...

対応パラメータ

すべてのプロバイダー = このモデルを提供するすべてのアップストリームが対応しています。一部のプロバイダー = リクエストを処理するアップストリームによって異なります。デフォルト = 未設定時に送信される値です。

パラメータプロバイダーデフォルト
frequency_penaltyすべてのプロバイダーデフォルトでは送信されません
include_reasoningすべてのプロバイダー-
logit_biasすべてのプロバイダー-
logprobsすべてのプロバイダー-
max_tokensすべてのプロバイダー-
min_pすべてのプロバイダー-
presence_penaltyすべてのプロバイダーデフォルトでは送信されません
reasoningすべてのプロバイダー-
reasoning_effortすべてのプロバイダー-
repetition_penaltyすべてのプロバイダーデフォルトでは送信されません
response_formatすべてのプロバイダー-
seedすべてのプロバイダー-
stopすべてのプロバイダー-
structured_outputsすべてのプロバイダー-
temperatureすべてのプロバイダー1
tool_choiceすべてのプロバイダー-
toolsすべてのプロバイダー-
top_kすべてのプロバイダーデフォルトでは送信されません
top_logprobsすべてのプロバイダー-
top_pすべてのプロバイダー0.95

よくあるご質問

類似モデル