D

DeepSeek: deepseek-v4-flash

deepseek-v4-flash
1M 컨텍스트384K 출력추론도구캐시구조화
출시일 Apr 24, 2026업데이트됨 Jun 18, 2026

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and high-throughput workloads, while maintaining strong reasoning and coding performance. The model includes hybrid attention for efficient long-context processing. Reasoning efforts high and xhigh are supported; xhigh maps to max reasoning. It is well suited for applications such as coding assistants, chat systems, and agent workflows where responsiveness and cost efficiency are important.

모드 chat토크나이저 DeepSeek사용 중단 Feb 2028deepseek-ai/DeepSeek-V4-Flash

요금

입력 가격
$0.06/ 100만 토큰$0.1455% 할인
출력 가격
$0.12/ 100만 토큰$0.2855% 할인
호환 엔드포인트 openai공급자 DeepSeek

가동 시간

성능

성능 데이터 로딩 중...

사용량 및 순위

사용량 불러오는 중...

지원 파라미터

모든 프로바이더 = 이 모델을 제공하는 모든 업스트림에서 지원됩니다. 일부 프로바이더 = 요청을 처리하는 업스트림에 따라 다릅니다. 기본값 = 설정하지 않았을 때 전송되는 값입니다.

파라미터프로바이더기본값
frequency_penalty모든 프로바이더-
include_reasoning모든 프로바이더-
logit_bias일부 프로바이더-
logprobs일부 프로바이더-
max_completion_tokens일부 프로바이더-
max_tokens모든 프로바이더-
min_p일부 프로바이더-
presence_penalty모든 프로바이더-
reasoning모든 프로바이더-
reasoning_effort모든 프로바이더-
repetition_penalty일부 프로바이더-
response_format모든 프로바이더-
seed모든 프로바이더-
stop모든 프로바이더-
structured_outputs모든 프로바이더-
temperature모든 프로바이더-
tool_choice모든 프로바이더-
tools모든 프로바이더-
top_a일부 프로바이더-
top_k모든 프로바이더-
top_logprobs일부 프로바이더-
top_p모든 프로바이더-

자주 묻는 질문

유사 모델