D

DeepSeek: deepseek-v4.1-flash:free

deepseek-v4.1-flash:free
1.0M contesto384K outputRagionamentoStrumentiVisioneCacheStrutturato
Rilasciato Sep 10, 2026Aggiornato Sep 15, 2026

DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on output from a 552B-parameter backbone, an asymmetric split that keeps per-token compute low relative to the model's total size. Image understanding is native to the architecture, with visual and text embeddings trained jointly from the start of pre-training rather than added afterward as in the earlier experimental V4 Flash Vision Exp. It is suited for coding, terminal, and computer-use agents, along with long-horizon tasks that must run to completion across many steps and long-context analysis. Compressed KV caching cuts cache memory to roughly a quarter of the previous Flash generation, significantly reducing costs on agentic workloads. DeepSeek positions it as the cost-efficient tier of the V4.1 family and reports that it exceeds V4 Pro on performance, speed, and task completion time.

Modalità chatTokenizer DeepSeekQuantizzazione fp8deepseek-ai/DeepSeek-V4.1-Flash

Prezzi

Prezzo di input
$0.00/ 1 M token
Prezzo di output
$0.00/ 1 M token
Endpoint compatibili openaiProvider DeepSeek

Disponibilità

Performance

Caricamento dati di performance...

Utilizzo e classifica

Caricamento utilizzo...

Parametri supportati

Tutti i provider = supportato da ogni upstream che serve questo modello. Alcuni provider = dipende dall'upstream che gestisce la richiesta. Predefinito = il valore inviato quando non lo imposti.

ParametroProviderPredefinito
frequency_penaltyTutti i provider-
include_reasoningTutti i provider-
logit_biasAlcuni provider-
logprobsAlcuni provider-
max_tokensTutti i provider-
min_pAlcuni provider-
presence_penaltyTutti i provider-
reasoningTutti i provider-
reasoning_effortTutti i provider-
repetition_penaltyTutti i provider-
response_formatTutti i provider-
seedAlcuni provider-
stopTutti i provider-
structured_outputsTutti i provider-
temperatureTutti i provider-
tool_choiceTutti i provider-
toolsTutti i provider-
top_kTutti i provider-
top_logprobsAlcuni provider-
top_pTutti i provider-

Domande frequenti

Modelli simili