Qwen: Qwen3.8 Omni Flash

Video input Image input Text input Audio input Text output
Author's Description

Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video understanding. It is suited for audio-video analysis and summarization,...

Key Specifications
Cost
$$$$
Context
1M
Released
Sep 20, 2026
Speed
Ability
Reliability
Supported Parameters

This model supports the following parameters:

Max Tokens Frequency Penalty Temperature Tool Choice Include Reasoning Reasoning Top P Stop Seed Response Format Logprobs Tools Presence Penalty Structured Outputs Top Logprobs
Features

This model supports the following features:

Tools Response Format Structured Outputs Reasoning
Performance Summary

Qwen3.8 Omni Flash, an omni-modal reasoning model from Alibaba, demonstrates a mixed performance profile with notable strengths in specific areas. The model tends to have longer response times, ranking in the 11th percentile for speed across benchmarks. However, it offers competitive pricing, placing in the 43rd percentile. Reliability is a significant strong point, with a 95% success rate indicating consistent and usable responses. In terms of accuracy, Qwen3.8 Omni Flash excels in Email Classification (99.0% accuracy, 86th percentile) and shows strong capabilities in Instruction Following (79.0% accuracy, 81st percentile). Its General Knowledge (96.0%) and Coding (86.9%) performances are solid, though not top-tier. A key weakness is observed in Reasoning, where it achieved only 56.0% accuracy (28th percentile), and in its ability to acknowledge uncertainty, with an 86.0% accuracy in Hallucinations (27th percentile), suggesting it may not always appropriately indicate when it lacks knowledge. Despite its agentic capabilities and native audio-video understanding, its overall speed remains a challenge.

Model Pricing

Current Pricing

Feature Price (per 1M tokens)
Prompt $0.15
Completion $0.47
Input Cache Read $0.016

Price History

Available Endpoints
Provider Endpoint Name Context Length Pricing (Input) Pricing (Output)
Alibaba
Alibaba | qwen/qwen3.8-omni-flash-20260918 1M $0.15 / 1M tokens $0.47 / 1M tokens
Benchmark Results
Benchmark Category Reasoning Strategy Free Executions Accuracy Cost Duration
Other Models by qwen