D

DeepSeek

21 個模型
DDeepSeek
新
deepseek-v4.1-flash

DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...

推理工具視覺快取
104.9萬$0.0079輸入 Token$0.1595% 優惠$0.03輸出 Token$0.6095% 優惠
DDeepSeek
deepseek-v4-pro-0813

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

推理工具快取
100萬$0.23輸入 Token$1.7487% 優惠$0.46輸出 Token$3.4887% 優惠
DDeepSeek
已棄用
deepseek-v4-flash-0731

DeepSeek-V4-Flash-0731 is the official release of DeepSeek-V4-Flash, superseding the preview version, with substantially enhanced agentic capabilities. It has the same model structure as...

推理工具視覺快取
100萬$0.04輸入 Token$0.1474% 優惠$0.07輸出 Token$0.2874% 優惠
DDeepSeek
已棄用
deepseek-v4-pro

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced...

推理工具快取
100萬$0.11輸入 Token$1.7494% 優惠$0.22輸出 Token$3.4894% 優惠
DDeepSeek
免費已棄用
deepseek-v4-pro:free

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced...

推理工具快取
100萬
DDeepSeek
已棄用
deepseek-v4-flash

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for...

推理工具視覺快取
100萬$0.0071輸入 Token$0.1495% 優惠$0.01輸出 Token$0.2895% 優惠
DDeepSeek
免費已棄用
deepseek-v4-flash:free

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for...

推理工具視覺快取
100萬
DDeepSeek
deepseek-v3.2

DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and agentic tool-use performance. It introduces DeepSeek Sparse Attention (DSA), a...

推理工具快取
12.8萬$0.56輸入 Token$0.583% 優惠$1.62輸出 Token$1.683% 優惠
DDeepSeek
已棄用
deepseek-v3.2-exp

DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures. It introduces DeepSeek Sparse Attention (DSA), a...

推理工具
16.4萬$0.56輸入 Token$0.583% 優惠$1.62輸出 Token$1.683% 優惠
DDeepSeek
deepseek-v3.1-terminus

DeepSeek-V3.1 Terminus is an update to DeepSeek V3.1 that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities,...

推理工具
16.4萬$0.54輸入 Token$2.00輸出 Token
DDeepSeek
deepseek-v3.1

DeepSeek-V3.1 Terminus is an update to DeepSeek V3.1 that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities,...

推理工具
16.4萬$0.02輸入 Token$0.2792% 優惠$0.08輸出 Token$1.0392% 優惠
DDeepSeek
deepseek-r1-0528

May 28th update to the original DeepSeek R1 Performance on par with OpenAI o1, but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference...

推理工具快取
12.8萬$1.12輸入 Token$3.0063% 優惠$4.48輸出 Token$12.0063% 優惠
DDeepSeek
deepseek-v3-0324

A state-of-the-art 671B MoE model (37B active) featuring Multi-head Latent Attention (MLA) and auxiliary-loss-free load balancing. It provides a hybrid "Think/Non-think" mode, delivering SOTA...

推理工具快取
12.8萬$0.50輸入 Token$1.87輸出 Token
DDeepSeek
deepseek-r1

DeepSeek R1 is here: Performance on par with OpenAI o1, but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass. Fully open-source...

推理工具
12.8萬$1.12輸入 Token$3.84輸出 Token
DDeepSeek
免費
deepseek-r1-distill-qwen-32b:free

DeepSeek R1 Distill Qwen 32B is a distilled large language model based on Qwen 2.5 32B, using outputs from DeepSeek R1. It outperforms OpenAI's o1-mini across various benchmarks, achieving new...

推理
8萬
DDeepSeek
已棄用
deepseek-v3

A state-of-the-art 671B MoE model (37B active) featuring Multi-head Latent Attention (MLA) and auxiliary-loss-free load balancing. It provides a hybrid "Think/Non-think" mode, delivering SOTA...

推理工具快取
12.8萬$0.56輸入 Token$1.1451% 優惠$2.24輸出 Token$4.5651% 優惠

目前繁忙

這些模型的所有供應商都已達到請求上限。限制解除後會自動恢復。