N

NVIDIA

15 个模型

当前繁忙

这些模型的所有供应商都已达到请求上限。限制解除后会自动恢复。

NNVIDIA
繁忙免费
nvidia-nemotron-3.5-lightning-30b-a3b:free

NVIDIA hybrid Mamba 2 and mixture of experts model, 30B total and 3B active parameters, with a 1M token context.

推理工具缓存
26.2万
NNVIDIA
繁忙免费
nemotron-3-embed-1b:free

NVIDIA Nemotron 3 Embed 1B is an open text embedding model from NVIDIA, optimized for high-throughput, low-latency retrieval. It is suited for enterprise search, RAG, code retrieval, and agentic...

26.2万
NNVIDIA
繁忙免费
nemotron-3-ultra-550b-a55b-free:free

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts...

推理工具缓存
104.9万
NNVIDIA
繁忙免费
nemotron-3-nano-omni-30b-a3b-reasoning-free:free

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and audio...

推理工具视觉
26.2万
NNVIDIA
繁忙免费
nemotron-3-super-120b-a12b-free:free

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid...

推理工具
26.2万
NNVIDIA
繁忙免费
nemotron-3-nano-30b-a3b:free

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully open with...

推理工具
26.2万
NNVIDIA
繁忙免费
llama-nemotron-embed-vl-1b-v2:free

The Llama Nemotron Embed VL 1B V2 embedding model is optimized for multimodal question-answering retrieval. The model can embed 'documents' in the form of image, text, or image and text combined....

视觉
1万
NNVIDIA
繁忙免费
mistral-nemotron:free
工具
12.8万