Qwen: Qwen3.8 Flash

Text input Image input Video input Text output
Author's Description

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

Key Specifications
Cost
$$$$
Context
1M
Parameters
3.8B (Rumoured)
Released
Aug 26, 2026
Speed
Ability
Reliability
Supported Parameters

This model supports the following parameters:

Temperature Include Reasoning Reasoning Presence Penalty Top Logprobs Stop Tool Choice Response Format Logprobs Top P Frequency Penalty Tools Max Tokens Structured Outputs Seed
Features

This model supports the following features:

Reasoning Tools Response Format Structured Outputs
Performance Summary

Qwen3.8 Flash, a multimodal reasoning model from Alibaba, demonstrates a balanced performance profile with notable strengths in reliability and specific task categories. While its speed performance is moderate, ranking in the 30th percentile, it offers competitive pricing, placing in the 49th percentile across benchmarks. A significant highlight is its strong reliability, boasting a 95% success rate, indicating consistent and usable responses with few technical issues. The model excels in Email Classification and Ethics, achieving 98.0% and 97.0% accuracy respectively, showcasing its proficiency in structured categorization and moral reasoning. Its Instruction Following capabilities are also strong at 77.0% accuracy, suggesting effectiveness in complex workflows. Qwen3.8 Flash performs well in Coding (90.0% accuracy) and Mathematics (92.8% accuracy), indicating solid technical and analytical skills. However, a key weakness is its performance in Hallucinations, with only 62.0% accuracy, suggesting a tendency to provide information rather than acknowledge uncertainty. Its General Knowledge accuracy (96.0%) is respectable but ranks lower compared to other models. The model's context length of 1,000,000 tokens supports extensive document and codebase analysis, aligning with its intended applications in agentic workflows and long-video analysis.

Model Pricing

Current Pricing

Feature Price (per 1M tokens)
Prompt $0.16
Completion $0.47
Input Cache Read $0.016
Input Cache Write $0.2

Price History

Available Endpoints
Provider Endpoint Name Context Length Pricing (Input) Pricing (Output)
Alibaba
Alibaba | qwen/qwen3.8-flash-20260826 1M $0.16 / 1M tokens $0.47 / 1M tokens
Benchmark Results
Benchmark Category Reasoning Strategy Free Executions Accuracy Cost Duration
Other Models by qwen