D
DeepSeek: deepseek-r1-distill-llama-70b:free
deepseek-r1-distill-llama-70b:free
131.1K konteks131.1K keluaranPenalaran
Dirilis Jan 23, 2025Batas pengetahuan Jul 2024Diperbarui Sep 14, 2026
DeepSeek R1 Distill Llama 70B is a distilled large language model based on Llama-3.3-70B-Instruct, using outputs from DeepSeek R1. The model combines advanced distillation techniques to achieve high performance across multiple benchmarks, including: AIME 2024 pass@1: 70.0 MATH-500 pass@1: 94.5 CodeForces Rating: 1633 The model leverages fine-tuning from DeepSeek R1's outputs, enabling competitive performance comparable to larger frontier models.
Harga
Harga input
$0.00/ 1 J token
Harga output
$0.00/ 1 J token
Jendela konteks 131.1K tokenEndpoint kompatibel openaiVendor DeepSeek
Waktu aktif
Performa
Memuat data performa...
Penggunaan & Peringkat
Memuat penggunaan...
Parameter didukung
Semua penyedia = didukung oleh setiap upstream yang melayani model ini. Sebagian penyedia = tergantung upstream yang menangani permintaan. Bawaan = nilai yang dikirim saat Anda tidak mengaturnya.
| Parameter | Penyedia | Default |
|---|---|---|
| frequency_penalty | Semua penyedia | - |
| include_reasoning | Semua penyedia | - |
| max_tokens | Semua penyedia | - |
| presence_penalty | Semua penyedia | - |
| reasoning | Semua penyedia | - |
| repetition_penalty | Semua penyedia | - |
| seed | Semua penyedia | - |
| stop | Semua penyedia | - |
| temperature | Semua penyedia | - |
| top_k | Semua penyedia | - |
| top_p | Semua penyedia | - |