D

DeepSeek: deepseek-v4-flash

deepseek-v4-flash
1M Kontext384K AusgabeReasoningToolsCacheStrukturiert
Veröffentlicht Apr 24, 2026Aktualisiert Jun 18, 2026

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and high-throughput workloads, while maintaining strong reasoning and coding performance. The model includes hybrid attention for efficient long-context processing. Reasoning efforts high and xhigh are supported; xhigh maps to max reasoning. It is well suited for applications such as coding assistants, chat systems, and agent workflows where responsiveness and cost efficiency are important.

Modus chatTokenizer DeepSeekAbkündigung Feb 2028deepseek-ai/DeepSeek-V4-Flash

Preise

Input-Preis
$0.06/ 1 Mio. Tokens$0.1455% Rabatt
Output-Preis
$0.12/ 1 Mio. Tokens$0.2855% Rabatt
Kompatible Endpunkte openaiAnbieter DeepSeek

Verfügbarkeit

Leistung

Lade Leistungsdaten...

Nutzung & Rang

Nutzung wird geladen...

Unterstützte Parameter

Alle Anbieter = jeder Upstream, der dieses Modell bereitstellt, unterstützt den Parameter. Einige Anbieter = hängt vom Upstream ab, der die Anfrage bearbeitet. Standard = der Wert, der gesendet wird, wenn nichts gesetzt ist.

ParameterAnbieterStandard
frequency_penaltyAlle Anbieter-
include_reasoningAlle Anbieter-
logit_biasEinige Anbieter-
logprobsEinige Anbieter-
max_completion_tokensEinige Anbieter-
max_tokensAlle Anbieter-
min_pEinige Anbieter-
presence_penaltyAlle Anbieter-
reasoningAlle Anbieter-
reasoning_effortAlle Anbieter-
repetition_penaltyEinige Anbieter-
response_formatAlle Anbieter-
seedAlle Anbieter-
stopAlle Anbieter-
structured_outputsAlle Anbieter-
temperatureAlle Anbieter-
tool_choiceAlle Anbieter-
toolsAlle Anbieter-
top_aEinige Anbieter-
top_kAlle Anbieter-
top_logprobsEinige Anbieter-
top_pAlle Anbieter-

Häufige Fragen

Ähnliche Modelle