N

NVIDIA: nemotron-3-ultra-550b-a55b-free:free

nemotron-3-ultra-550b-a55b-free:free
104.9万 コンテキスト3.3万 出力推論ツールキャッシュ構造化
リリース日 Jun 4, 2026更新日 Sep 4, 2026

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it supports text input and output with a context window of up to 1M tokens. It is suited for long-running agentic workflows, including agent orchestration, coding agents, deep research, and complex enterprise tasks. It is particularly strong at multi-step reasoning and planning, with high-throughput inference designed for high-volume agent pipelines. It is part of the NVIDIA Nemotron family of open models for agentic AI.

モード chatトークナイザー Other量子化 fp4nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16

このモデルのプロバイダーは現在すべて混雑しています

すべての上流プロバイダーがレート制限に達しました。制限が解除されるとモデルは自動的に復帰します。通常は数時間以内です。しばらくしてから再試行するか、別のモデルに切り替えてください。

このモデルをDiscordでリクエスト

料金

入力料金
$0.00/ 100万トークン
出力料金
$0.00/ 100万トークン
対応エンドポイント openaiベンダー NVIDIA

稼働率

パフォーマンス

パフォーマンスデータを読み込み中...

使用状況とランキング

使用状況を読み込み中...

対応パラメータ

すべてのプロバイダー = このモデルを提供するすべてのアップストリームが対応しています。一部のプロバイダー = リクエストを処理するアップストリームによって異なります。デフォルト = 未設定時に送信される値です。

パラメータプロバイダーデフォルト
frequency_penaltyすべてのプロバイダーデフォルトでは送信されません
include_reasoningすべてのプロバイダー-
logit_bias一部のプロバイダー-
max_tokensすべてのプロバイダー-
min_p一部のプロバイダー-
presence_penaltyすべてのプロバイダーデフォルトでは送信されません
reasoningすべてのプロバイダー-
reasoning_effortすべてのプロバイダー-
repetition_penalty一部のプロバイダーデフォルトでは送信されません
response_formatすべてのプロバイダー-
seed一部のプロバイダー-
stopすべてのプロバイダー-
structured_outputs一部のプロバイダー-
temperatureすべてのプロバイダー1
tool_choiceすべてのプロバイダー-
toolsすべてのプロバイダー-
top_kすべてのプロバイダーデフォルトでは送信されません
top_pすべてのプロバイダー0.95

よくあるご質問

類似モデル