NVIDIA: nemotron-3-ultra-550b-a55b:free
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it supports text input and output with a context window of up to 1M tokens. It is suited for long-running agentic workflows, including agent orchestration, coding agents, deep research, and complex enterprise tasks. It is particularly strong at multi-step reasoning and planning, with high-throughput inference designed for high-volume agent pipelines. It is part of the NVIDIA Nemotron family of open models for agentic AI.
Prezzi
Disponibilità
Performance
Utilizzo e classifica
Parametri supportati
Tutti i provider = supportato da ogni upstream che serve questo modello. Alcuni provider = dipende dall'upstream che gestisce la richiesta. Predefinito = il valore inviato quando non lo imposti.
| Parametro | Provider | Predefinito |
|---|---|---|
| frequency_penalty | Tutti i provider | Non inviato per impostazione predefinita |
| include_reasoning | Tutti i provider | - |
| logit_bias | Tutti i provider | - |
| max_tokens | Tutti i provider | - |
| min_p | Tutti i provider | - |
| presence_penalty | Tutti i provider | Non inviato per impostazione predefinita |
| reasoning | Tutti i provider | - |
| reasoning_effort | Tutti i provider | - |
| repetition_penalty | Tutti i provider | Non inviato per impostazione predefinita |
| response_format | Tutti i provider | - |
| seed | Alcuni provider | - |
| stop | Tutti i provider | - |
| structured_outputs | Tutti i provider | - |
| temperature | Tutti i provider | 1 |
| tool_choice | Tutti i provider | - |
| tools | Tutti i provider | - |
| top_k | Tutti i provider | Non inviato per impostazione predefinita |
| top_p | Tutti i provider | 0.95 |