D

DeepSeek: deepseek-v4-flash-0731:free

deepseek-v4-flash-0731:free
1M contesto384K outputRagionamentoStrumentiStrumenti paralleliVisioneCacheStrutturato
Rilasciato Jul 31, 2026Aggiornato Oct 8, 2026

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and high-throughput workloads, while maintaining strong reasoning and coding performance. The model includes hybrid attention for efficient long-context processing. Reasoning efforts high and xhigh are supported; xhigh maps to max reasoning. It is well suited for applications such as coding assistants, chat systems, and agent workflows where responsiveness and cost efficiency are important.

Modalità chatTokenizer DeepSeekQuantizzazione fp8Deprecazione Dec 2026deepseek-ai/DeepSeek-V4-Flash-0731

Tutti i provider di questo modello sono occupati al momento

Ogni provider a monte ha raggiunto il suo limite di velocità. Il modello torna automaticamente quando i limiti si allentano, di solito entro poche ore. Riprova tra poco o passa a un altro modello.

Richiedi questo modello su Discord

Prezzi

Prezzo di input
$0.00/ 1 M token
Prezzo di output
$0.00/ 1 M token
Endpoint compatibili openaiProvider DeepSeek

Disponibilità

Performance

Caricamento dati di performance...

Utilizzo e classifica

Caricamento utilizzo...

Parametri supportati

Tutti i provider = supportato da ogni upstream che serve questo modello. Alcuni provider = dipende dall'upstream che gestisce la richiesta. Predefinito = il valore inviato quando non lo imposti.

ParametroProviderPredefinito
frequency_penaltyTutti i provider-
include_reasoningTutti i provider-
logit_biasAlcuni provider-
logprobsTutti i provider-
max_tokensTutti i provider-
min_pAlcuni provider-
parallel_tool_callsAlcuni provider-
presence_penaltyTutti i provider-
reasoningTutti i provider-
reasoning_effortTutti i provider-
repetition_penaltyTutti i provider-
response_formatTutti i provider-
seedTutti i provider-
stopTutti i provider-
structured_outputsTutti i provider-
temperatureTutti i provider-
tool_choiceTutti i provider-
toolsTutti i provider-
top_aAlcuni provider-
top_kTutti i provider-
top_logprobsTutti i provider-
top_pTutti i provider-

Domande frequenti

Modelli simili