D

DeepSeek: deepseek-v4.1-flash:free

deepseek-v4.1-flash:free
104.9万 上下文38.4万 输出推理工具视觉缓存结构化
发布日期 Sep 10, 2026更新时间 Sep 15, 2026

DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on output from a 552B-parameter backbone, an asymmetric split that keeps per-token compute low relative to the model's total size. Image understanding is native to the architecture, with visual and text embeddings trained jointly from the start of pre-training rather than added afterward as in the earlier experimental V4 Flash Vision Exp. It is suited for coding, terminal, and computer-use agents, along with long-horizon tasks that must run to completion across many steps and long-context analysis. Compressed KV caching cuts cache memory to roughly a quarter of the previous Flash generation, significantly reducing costs on agentic workloads. DeepSeek positions it as the cost-efficient tier of the V4.1 family and reports that it exceeds V4 Pro on performance, speed, and task completion time.

模式 chat分词器 DeepSeek量化 fp8deepseek-ai/DeepSeek-V4.1-Flash

该模型的所有提供方目前都很繁忙

每个上游提供方都已达到速率上限。上限解除后模型会自动恢复,通常在数小时内。请稍后重试,或切换到其他模型。

在 Discord 上申请此模型

价格

输入价格
$0.00/ 100 万 token
输出价格
$0.00/ 100 万 token
兼容端点 -供应商 DeepSeek

在线时长

性能

正在加载性能数据...

使用量与排名

正在加载使用量……

支持的参数

所有提供方:为该模型提供服务的每个上游都支持。部分提供方:取决于处理请求的上游。默认值:未设置时发送的值。

参数提供方默认值
frequency_penalty所有提供方-
include_reasoning所有提供方-
logit_bias部分提供方-
logprobs部分提供方-
max_tokens所有提供方-
min_p部分提供方-
presence_penalty所有提供方-
reasoning所有提供方-
reasoning_effort所有提供方-
repetition_penalty所有提供方-
response_format所有提供方-
seed部分提供方-
stop所有提供方-
structured_outputs所有提供方-
temperature所有提供方-
tool_choice所有提供方-
tools所有提供方-
top_k所有提供方-
top_logprobs部分提供方-
top_p所有提供方-

常见问题

相似模型