N

NVIDIA: llama-3.3-nemotron-super-49b-v1:free

llama-3.3-nemotron-super-49b-v1:free
131.1K 컨텍스트131.1K 출력도구
출시일 Mar 18, 2025업데이트됨 Jun 18, 2026

Llama-3.3-Nemotron-Super-49B-v1 is a large language model (LLM) optimized for advanced reasoning, conversational interactions, retrieval-augmented generation (RAG), and tool-calling tasks. Derived from Meta's Llama-3.3-70B-Instruct, it employs a Neural Architecture Search (NAS) approach, significantly enhancing efficiency and reducing memory requirements. This allows the model to support a context length of up to 128K tokens and fit efficiently on single high-performance GPUs, such as NVIDIA H200. Note: you must include detailed thinking on in the system prompt to enable reasoning. Please see Usage Recommendations for more.

모드 chat

요금

입력 가격
$0.00/ 100만 토큰
출력 가격
$0.00/ 100만 토큰
컨텍스트 윈도우 131.1K 토큰호환 엔드포인트 openai공급자 NVIDIA

가동 시간

성능

성능 데이터 로딩 중...

사용량 및 순위

사용량 불러오는 중...

자주 묻는 질문

유사 모델