N
NVIDIA: nemotron-3-embed-1b:free
nemotron-3-embed-1b:free
262.1K contextembedding
Released Jul 14, 2026Updated Sep 21, 2026
NVIDIA Nemotron 3 Embed 1B is an open text embedding model from NVIDIA, optimized for high-throughput, low-latency retrieval. It is suited for enterprise search, RAG, code retrieval, and agentic retrieval workflows, retaining more than 95% of the 8B model’s accuracy in a smaller deployment footprint.
Mode embedding
All providers for this model are busy right now
Every upstream provider has hit its rate limit. The model comes back automatically once limits lift, usually within hours. Try again in a little while or switch to another model.
Request this model on DiscordPricing
Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Context window 262.1K tokensCompatible endpoints openaiVendor NVIDIA
Uptime
Performance
Loading performance data...
Usage & Ranking
Loading usage...