D

DeepSeek: deepseek-v4-flash-0731:free

deepseek-v4-flash-0731:free
100万 コンテキスト38.4万 出力推論ツール並列ツールビジョンキャッシュ構造化
リリース日 Jul 31, 2026更新日 Oct 8, 2026

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and high-throughput workloads, while maintaining strong reasoning and coding performance. The model includes hybrid attention for efficient long-context processing. Reasoning efforts high and xhigh are supported; xhigh maps to max reasoning. It is well suited for applications such as coding assistants, chat systems, and agent workflows where responsiveness and cost efficiency are important.

モード chatトークナイザー DeepSeek量子化 fp8非推奨 Dec 2026deepseek-ai/DeepSeek-V4-Flash-0731

このモデルのプロバイダーは現在すべて混雑しています

すべての上流プロバイダーがレート制限に達しました。制限が解除されるとモデルは自動的に復帰します。通常は数時間以内です。しばらくしてから再試行するか、別のモデルに切り替えてください。

このモデルをDiscordでリクエスト

料金

入力料金
$0.00/ 100万トークン
出力料金
$0.00/ 100万トークン
対応エンドポイント openaiベンダー DeepSeek

稼働率

パフォーマンス

パフォーマンスデータを読み込み中...

使用状況とランキング

使用状況を読み込み中...

対応パラメータ

すべてのプロバイダー = このモデルを提供するすべてのアップストリームが対応しています。一部のプロバイダー = リクエストを処理するアップストリームによって異なります。デフォルト = 未設定時に送信される値です。

パラメータプロバイダーデフォルト
frequency_penaltyすべてのプロバイダー-
include_reasoningすべてのプロバイダー-
logit_bias一部のプロバイダー-
logprobsすべてのプロバイダー-
max_tokensすべてのプロバイダー-
min_p一部のプロバイダー-
parallel_tool_calls一部のプロバイダー-
presence_penaltyすべてのプロバイダー-
reasoningすべてのプロバイダー-
reasoning_effortすべてのプロバイダー-
repetition_penaltyすべてのプロバイダー-
response_formatすべてのプロバイダー-
seedすべてのプロバイダー-
stopすべてのプロバイダー-
structured_outputsすべてのプロバイダー-
temperatureすべてのプロバイダー-
tool_choiceすべてのプロバイダー-
toolsすべてのプロバイダー-
top_a一部のプロバイダー-
top_kすべてのプロバイダー-
top_logprobsすべてのプロバイダー-
top_pすべてのプロバイダー-

よくあるご質問

類似モデル