A
Alibaba: qwen3.8-omni-flash
qwen3.8-omni-flash
1M context131.1K outReasoningToolsVisionAudio inVideoCacheStructuredWeb search
Released Sep 21, 2026Updated Sep 29, 2026
Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video understanding. It is suited for audio-video analysis and summarization, video editing and production workflows, audio-video dialogue, coding, knowledge work, and GUI interaction, and it is particularly strong at long-form multimedia tasks that combine speech, sound, and visual context. It also supports two-channel and four-channel spatial audio understanding.
Mode chatTokenizer Qwen
Pricing
Input price
$0.30/ 1M tokens
Output price
$0.94/ 1M tokens
Compatible endpoints openaiVendor Alibaba
Uptime
Performance
Loading performance data...
Usage & Ranking
Loading usage...
Supported parameters
All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.
| Parameter | Providers | Default |
|---|---|---|
| frequency_penalty | All providers | - |
| include_reasoning | All providers | - |
| logprobs | All providers | - |
| max_tokens | All providers | - |
| presence_penalty | All providers | - |
| reasoning | All providers | - |
| response_format | All providers | - |
| seed | All providers | - |
| stop | All providers | - |
| structured_outputs | All providers | - |
| temperature | All providers | - |
| tool_choice | All providers | - |
| tools | All providers | - |
| top_k | All providers | - |
| top_logprobs | All providers | - |
| top_p | All providers | - |