D

DeepSeek

21 modelli
DDeepSeek
Nuovo
deepseek-v4.1-flash

DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...

RagionamentoStrumentiVisioneCache
1.0M$0.0079token di input$0.1595% di sconto$0.03token di output$0.6095% di sconto
DDeepSeek
deepseek-v4-pro-0813

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

RagionamentoStrumentiCache
1M$0.23token di input$1.7487% di sconto$0.46token di output$3.4887% di sconto
DDeepSeek
Deprecato
deepseek-v4-flash-0731

DeepSeek-V4-Flash-0731 is the official release of DeepSeek-V4-Flash, superseding the preview version, with substantially enhanced agentic capabilities. It has the same model structure as...

RagionamentoStrumentiVisioneCache
1M$0.04token di input$0.1474% di sconto$0.07token di output$0.2874% di sconto
DDeepSeek
Deprecato
deepseek-v4-pro

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced...

RagionamentoStrumentiCache
1M$0.11token di input$1.7494% di sconto$0.22token di output$3.4894% di sconto
DDeepSeek
gratisDeprecato
deepseek-v4-pro:free

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced...

RagionamentoStrumentiCache
1M
DDeepSeek
Deprecato
deepseek-v4-flash

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for...

RagionamentoStrumentiVisioneCache
1M$0.0071token di input$0.1495% di sconto$0.01token di output$0.2895% di sconto
DDeepSeek
gratisDeprecato
deepseek-v4-flash:free

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for...

RagionamentoStrumentiVisioneCache
1M
DDeepSeek
deepseek-v3.2

DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and agentic tool-use performance. It introduces DeepSeek Sparse Attention (DSA), a...

RagionamentoStrumentiCache
128K$0.56token di input$0.583% di sconto$1.62token di output$1.683% di sconto
DDeepSeek
Deprecato
deepseek-v3.2-exp

DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures. It introduces DeepSeek Sparse Attention (DSA), a...

RagionamentoStrumenti
163.8K$0.56token di input$0.583% di sconto$1.62token di output$1.683% di sconto
DDeepSeek
deepseek-v3.1-terminus

DeepSeek-V3.1 Terminus is an update to DeepSeek V3.1 that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities,...

RagionamentoStrumenti
163.8K$0.54token di input$2.00token di output
DDeepSeek
deepseek-v3.1

DeepSeek-V3.1 Terminus is an update to DeepSeek V3.1 that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities,...

RagionamentoStrumenti
163.8K$0.02token di input$0.2792% di sconto$0.08token di output$1.0392% di sconto
DDeepSeek
deepseek-r1-0528

May 28th update to the original DeepSeek R1 Performance on par with OpenAI o1, but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference...

RagionamentoStrumentiCache
128K$1.12token di input$3.0063% di sconto$4.48token di output$12.0063% di sconto
DDeepSeek
deepseek-v3-0324

A state-of-the-art 671B MoE model (37B active) featuring Multi-head Latent Attention (MLA) and auxiliary-loss-free load balancing. It provides a hybrid "Think/Non-think" mode, delivering SOTA...

RagionamentoStrumentiCache
128K$0.50token di input$1.87token di output
DDeepSeek
deepseek-r1

DeepSeek R1 is here: Performance on par with OpenAI o1, but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass. Fully open-source...

RagionamentoStrumenti
128K$1.12token di input$3.84token di output
DDeepSeek
gratis
deepseek-r1-distill-qwen-32b:free

DeepSeek R1 Distill Qwen 32B is a distilled large language model based on Qwen 2.5 32B, using outputs from DeepSeek R1. It outperforms OpenAI's o1-mini across various benchmarks, achieving new...

Ragionamento
80K
DDeepSeek
Deprecato
deepseek-v3

A state-of-the-art 671B MoE model (37B active) featuring Multi-head Latent Attention (MLA) and auxiliary-loss-free load balancing. It provides a hybrid "Think/Non-think" mode, delivering SOTA...

RagionamentoStrumentiCache
128K$0.56token di input$1.1451% di sconto$2.24token di output$4.5651% di sconto

Al momento occupati

Tutti i provider di questi modelli hanno raggiunto il limite di richieste. Tornano disponibili automaticamente quando i limiti si azzerano.