D

DeepSeek

21 models
DDeepSeek
New
deepseek-v4.1-flash

DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...

ReasoningToolsVisionCache
1.0M$0.0079input tokens$0.1595% off$0.03output tokens$0.6095% off
DDeepSeek
deepseek-v4-pro-0813

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

ReasoningToolsCache
1M$0.23input tokens$1.7487% off$0.46output tokens$3.4887% off
DDeepSeek
Deprecated
deepseek-v4-flash-0731

DeepSeek-V4-Flash-0731 is the official release of DeepSeek-V4-Flash, superseding the preview version, with substantially enhanced agentic capabilities. It has the same model structure as...

ReasoningToolsVisionCache
1M$0.04input tokens$0.1474% off$0.07output tokens$0.2874% off
DDeepSeek
Deprecated
deepseek-v4-pro

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced...

ReasoningToolsCache
1M$0.11input tokens$1.7494% off$0.22output tokens$3.4894% off
DDeepSeek
freeDeprecated
deepseek-v4-pro:free

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced...

ReasoningToolsCache
1M
DDeepSeek
Deprecated
deepseek-v4-flash

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for...

ReasoningToolsVisionCache
1M$0.0071input tokens$0.1495% off$0.01output tokens$0.2895% off
DDeepSeek
freeDeprecated
deepseek-v4-flash:free

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for...

ReasoningToolsVisionCache
1M
DDeepSeek
deepseek-v3.2

DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and agentic tool-use performance. It introduces DeepSeek Sparse Attention (DSA), a...

ReasoningToolsCache
128K$0.56input tokens$0.583% off$1.62output tokens$1.683% off
DDeepSeek
Deprecated
deepseek-v3.2-exp

DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures. It introduces DeepSeek Sparse Attention (DSA), a...

ReasoningTools
163.8K$0.56input tokens$0.583% off$1.62output tokens$1.683% off
DDeepSeek
deepseek-v3.1-terminus

DeepSeek-V3.1 Terminus is an update to DeepSeek V3.1 that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities,...

ReasoningTools
163.8K$0.54input tokens$2.00output tokens
DDeepSeek
deepseek-v3.1

DeepSeek-V3.1 Terminus is an update to DeepSeek V3.1 that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities,...

ReasoningTools
163.8K$0.02input tokens$0.2792% off$0.08output tokens$1.0392% off
DDeepSeek
deepseek-r1-0528

May 28th update to the original DeepSeek R1 Performance on par with OpenAI o1, but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference...

ReasoningToolsCache
128K$1.12input tokens$3.0063% off$4.48output tokens$12.0063% off
DDeepSeek
deepseek-v3-0324

A state-of-the-art 671B MoE model (37B active) featuring Multi-head Latent Attention (MLA) and auxiliary-loss-free load balancing. It provides a hybrid "Think/Non-think" mode, delivering SOTA...

ReasoningToolsCache
128K$0.50input tokens$1.87output tokens
DDeepSeek
deepseek-r1

DeepSeek R1 is here: Performance on par with OpenAI o1, but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass. Fully open-source...

ReasoningTools
128K$1.12input tokens$3.84output tokens
DDeepSeek
free
deepseek-r1-distill-qwen-32b:free

DeepSeek R1 Distill Qwen 32B is a distilled large language model based on Qwen 2.5 32B, using outputs from DeepSeek R1. It outperforms OpenAI's o1-mini across various benchmarks, achieving new...

Reasoning
80K
DDeepSeek
Deprecated
deepseek-v3

A state-of-the-art 671B MoE model (37B active) featuring Multi-head Latent Attention (MLA) and auxiliary-loss-free load balancing. It provides a hybrid "Think/Non-think" mode, delivering SOTA...

ReasoningToolsCache
128K$0.56input tokens$1.1451% off$2.24output tokens$4.5651% off

Currently busy

Every provider for these models has hit its rate limit. They come back automatically once the limits lift.