D

DeepSeek

21 מודלים
DDeepSeek
חדש
deepseek-v4.1-flash

DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...

חשיבהכליםראייהמטמון
1.0M$0.01טוקני קלט$0.1590% הנחה$0.06טוקני פלט$0.6090% הנחה
DDeepSeek
deepseek-v4-pro-0813

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

חשיבהכליםמטמון
1M$0.23טוקני קלט$1.7487% הנחה$0.46טוקני פלט$3.4887% הנחה
DDeepSeek
הוצא משימוש
deepseek-v4-flash-0731

DeepSeek-V4-Flash-0731 is the official release of DeepSeek-V4-Flash, superseding the preview version, with substantially enhanced agentic capabilities. It has the same model structure as...

חשיבהכליםראייהמטמון
1M$0.04טוקני קלט$0.1474% הנחה$0.07טוקני פלט$0.2874% הנחה
DDeepSeek
הוצא משימוש
deepseek-v4-pro

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced...

חשיבהכליםמטמון
1M$0.11טוקני קלט$1.7494% הנחה$0.22טוקני פלט$3.4894% הנחה
DDeepSeek
חינםהוצא משימוש
deepseek-v4-pro:free

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced...

חשיבהכליםמטמון
1M
DDeepSeek
הוצא משימוש
deepseek-v4-flash

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for...

חשיבהכליםראייהמטמון
1M$0.0071טוקני קלט$0.1495% הנחה$0.01טוקני פלט$0.2895% הנחה
DDeepSeek
חינםהוצא משימוש
deepseek-v4-flash:free

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for...

חשיבהכליםראייהמטמון
1M
DDeepSeek
deepseek-v3.2

DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and agentic tool-use performance. It introduces DeepSeek Sparse Attention (DSA), a...

חשיבהכליםמטמון
128K$0.56טוקני קלט$0.583% הנחה$1.62טוקני פלט$1.683% הנחה
DDeepSeek
הוצא משימוש
deepseek-v3.2-exp

DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures. It introduces DeepSeek Sparse Attention (DSA), a...

חשיבהכלים
163.8K$0.56טוקני קלט$0.583% הנחה$1.62טוקני פלט$1.683% הנחה
DDeepSeek
deepseek-v3.1-terminus

DeepSeek-V3.1 Terminus is an update to DeepSeek V3.1 that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities,...

חשיבהכלים
163.8K$0.54טוקני קלט$2.00טוקני פלט
DDeepSeek
deepseek-v3.1

DeepSeek-V3.1 Terminus is an update to DeepSeek V3.1 that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities,...

חשיבהכלים
163.8K$0.02טוקני קלט$0.2792% הנחה$0.08טוקני פלט$1.0392% הנחה
DDeepSeek
deepseek-r1-0528

May 28th update to the original DeepSeek R1 Performance on par with OpenAI o1, but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference...

חשיבהכליםמטמון
128K$1.12טוקני קלט$3.0063% הנחה$4.48טוקני פלט$12.0063% הנחה
DDeepSeek
deepseek-v3-0324

A state-of-the-art 671B MoE model (37B active) featuring Multi-head Latent Attention (MLA) and auxiliary-loss-free load balancing. It provides a hybrid "Think/Non-think" mode, delivering SOTA...

חשיבהכליםמטמון
128K$0.50טוקני קלט$1.87טוקני פלט
DDeepSeek
deepseek-r1

DeepSeek R1 is here: Performance on par with OpenAI o1, but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass. Fully open-source...

חשיבהכלים
128K$1.12טוקני קלט$3.84טוקני פלט
DDeepSeek
חינם
deepseek-r1-distill-qwen-32b:free

DeepSeek R1 Distill Qwen 32B is a distilled large language model based on Qwen 2.5 32B, using outputs from DeepSeek R1. It outperforms OpenAI's o1-mini across various benchmarks, achieving new...

חשיבה
80K
DDeepSeek
הוצא משימוש
deepseek-v3

A state-of-the-art 671B MoE model (37B active) featuring Multi-head Latent Attention (MLA) and auxiliary-loss-free load balancing. It provides a hybrid "Think/Non-think" mode, delivering SOTA...

חשיבהכליםמטמון
128K$0.56טוקני קלט$1.1451% הנחה$2.24טוקני פלט$4.5651% הנחה

עמוסים כרגע

כל הספקים של הדגמים האלה הגיעו למגבלת הבקשות. הם חוזרים אוטומטית ברגע שהמגבלות מתאפסות.