N

NVIDIA: nemotron-3-nano-omni-30b-a3b-reasoning-free:free

nemotron-3-nano-omni-30b-a3b-reasoning-free:free
262.1K contextReasoningToolsVision
Released Apr 28, 2026Updated Sep 4, 2026

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and audio inputs and produces text output, enabling agents to perceive and reason across modalities in a single inference loop. Built on a hybrid MoE Transformer-Mamba architecture with Conv3D video layers and Efficient Video Sampling (EVS), it delivers approximately 2× higher throughput and 2.5× lower compute for video reasoning versus separate vision + speech pipelines. It supports up to 300K context length and a 16,384 reasoning budget, with extended thinking enabled via reasoning.enabled on OpenRouter.

All providers for this model are busy right now

Every upstream provider has hit its rate limit. The model comes back automatically once limits lift, usually within hours. Try again in a little while or switch to another model.

Request this model on Discord

Pricing

Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Context window 262.1K tokensCompatible endpoints openaiVendor NVIDIA

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Frequently asked questions

Similar models