D

DeepSeek: deepseek-v4.1-flash:free

deepseek-v4.1-flash:free
104.9萬 上下文38.4萬 輸出推理工具視覺快取結構化
發布日期 Sep 10, 2026更新時間 Sep 15, 2026

DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on output from a 552B-parameter backbone, an asymmetric split that keeps per-token compute low relative to the model's total size. Image understanding is native to the architecture, with visual and text embeddings trained jointly from the start of pre-training rather than added afterward as in the earlier experimental V4 Flash Vision Exp. It is suited for coding, terminal, and computer-use agents, along with long-horizon tasks that must run to completion across many steps and long-context analysis. Compressed KV caching cuts cache memory to roughly a quarter of the previous Flash generation, significantly reducing costs on agentic workloads. DeepSeek positions it as the cost-efficient tier of the V4.1 family and reports that it exceeds V4 Pro on performance, speed, and task completion time.

模式 chat分詞器 DeepSeek量化 fp8deepseek-ai/DeepSeek-V4.1-Flash

該模型的所有供應商目前都很繁忙

每個上游供應商都已達到速率上限。上限解除後模型會自動恢復,通常在數小時內。請稍後重試,或改用其他模型。

在 Discord 上申請此模型

價格

輸入價格
$0.00/ 100 萬 token
輸出價格
$0.00/ 100 萬 token
相容端點 -供應商 DeepSeek

在線時長

效能

正在載入效能資料...

使用量與排名

正在載入使用量……

支援的參數

所有提供方:為該模型提供服務的每個上游都支援。部分提供方:取決於處理請求的上游。預設值:未設定時傳送的值。

參數提供方預設值
frequency_penalty所有提供方-
include_reasoning所有提供方-
logit_bias部分提供方-
logprobs部分提供方-
max_tokens所有提供方-
min_p部分提供方-
presence_penalty所有提供方-
reasoning所有提供方-
reasoning_effort所有提供方-
repetition_penalty所有提供方-
response_format所有提供方-
seed部分提供方-
stop所有提供方-
structured_outputs所有提供方-
temperature所有提供方-
tool_choice所有提供方-
tools所有提供方-
top_k所有提供方-
top_logprobs部分提供方-
top_p所有提供方-

常見問題

相似模型