A

Alibaba: qwen3.8-omni-flash

qwen3.8-omni-flash
1M context131.1K outReasoningToolsVisionAudio inVideoCacheStructuredWeb search
Released Sep 21, 2026Updated Sep 29, 2026

Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video understanding. It is suited for audio-video analysis and summarization, video editing and production workflows, audio-video dialogue, coding, knowledge work, and GUI interaction, and it is particularly strong at long-form multimedia tasks that combine speech, sound, and visual context. It also supports two-channel and four-channel spatial audio understanding.

Mode chatTokenizer Qwen

Pricing

Input price
$0.30/ 1M tokens
Output price
$0.94/ 1M tokens
Compatible endpoints openaiVendor Alibaba

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Supported parameters

All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.

ParameterProvidersDefault
frequency_penaltyAll providers-
include_reasoningAll providers-
logprobsAll providers-
max_tokensAll providers-
presence_penaltyAll providers-
reasoningAll providers-
response_formatAll providers-
seedAll providers-
stopAll providers-
structured_outputsAll providers-
temperatureAll providers-
tool_choiceAll providers-
toolsAll providers-
top_kAll providers-
top_logprobsAll providers-
top_pAll providers-

Frequently asked questions

Similar models