D

DeepSeek: deepseek-r1-distill-llama-70b:free

deepseek-r1-distill-llama-70b:free
131.1K contesto131.1K outputRagionamento
Rilasciato Jan 23, 2025Limite delle conoscenze Jul 2024Aggiornato Sep 14, 2026

DeepSeek R1 Distill Llama 70B is a distilled large language model based on Llama-3.3-70B-Instruct, using outputs from DeepSeek R1. The model combines advanced distillation techniques to achieve high performance across multiple benchmarks, including: AIME 2024 pass@1: 70.0 MATH-500 pass@1: 94.5 CodeForces Rating: 1633 The model leverages fine-tuning from DeepSeek R1's outputs, enabling competitive performance comparable to larger frontier models.

Modalità chatTokenizer Llama3Quantizzazione bf16deepseek-ai/DeepSeek-R1-Distill-Llama-70B

Prezzi

Prezzo di input
$0.00/ 1 M token
Prezzo di output
$0.00/ 1 M token
Finestra di contesto 131.1K tokenEndpoint compatibili openaiProvider DeepSeek

Disponibilità

Performance

Caricamento dati di performance...

Utilizzo e classifica

Caricamento utilizzo...

Parametri supportati

Tutti i provider = supportato da ogni upstream che serve questo modello. Alcuni provider = dipende dall'upstream che gestisce la richiesta. Predefinito = il valore inviato quando non lo imposti.

ParametroProviderPredefinito
frequency_penaltyTutti i provider-
include_reasoningTutti i provider-
max_tokensTutti i provider-
presence_penaltyTutti i provider-
reasoningTutti i provider-
repetition_penaltyTutti i provider-
seedTutti i provider-
stopTutti i provider-
temperatureTutti i provider-
top_kTutti i provider-
top_pTutti i provider-

Domande frequenti

Modelli simili