N

NVIDIA: nemotron-3-ultra-550b-a55b:free

nemotron-3-ultra-550b-a55b:free
26.2万 コンテキスト3.3万 出力推論ツールキャッシュ構造化
リリース日 Jun 4, 2026更新日 Jun 18, 2026

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it supports text input and output with a context window of up to 1M tokens. It is suited for long-running agentic workflows, including agent orchestration, coding agents, deep research, and complex enterprise tasks. It is particularly strong at multi-step reasoning and planning, with high-throughput inference designed for high-volume agent pipelines. It is part of the NVIDIA Nemotron family of open models for agentic AI.

モード chatトークナイザー Othernvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16

料金

入力料金
$0.00/ 100万トークン
出力料金
$0.00/ 100万トークン
コンテキストウィンドウ 512.3K トークン対応エンドポイント openaiベンダー NVIDIA

稼働率

パフォーマンス

パフォーマンスデータを読み込み中...

使用状況とランキング

使用状況を読み込み中...

対応パラメータ

すべてのプロバイダー = このモデルを提供するすべてのアップストリームが対応しています。一部のプロバイダー = リクエストを処理するアップストリームによって異なります。デフォルト = 未設定時に送信される値です。

パラメータプロバイダーデフォルト
frequency_penaltyすべてのプロバイダーデフォルトでは送信されません
include_reasoningすべてのプロバイダー-
logit_biasすべてのプロバイダー-
max_tokensすべてのプロバイダー-
min_pすべてのプロバイダー-
presence_penaltyすべてのプロバイダーデフォルトでは送信されません
reasoningすべてのプロバイダー-
reasoning_effortすべてのプロバイダー-
repetition_penaltyすべてのプロバイダーデフォルトでは送信されません
response_formatすべてのプロバイダー-
seed一部のプロバイダー-
stopすべてのプロバイダー-
structured_outputsすべてのプロバイダー-
temperatureすべてのプロバイダー1
tool_choiceすべてのプロバイダー-
toolsすべてのプロバイダー-
top_kすべてのプロバイダーデフォルトでは送信されません
top_pすべてのプロバイダー0.95

よくあるご質問

類似モデル