Qwen: Qwen3.7 Flash

Text input Image input Video input Text output
Author's Description

Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...

Key Specifications
Cost
$$
Context
1M
Parameters
3.7B (Rumoured)
Released
Jul 27, 2026
Speed
Ability
Reliability
Supported Parameters

This model supports the following parameters:

Top P Top Logprobs Logprobs Tool Choice Tools Response Format Include Reasoning Temperature Presence Penalty Reasoning Max Tokens Seed
Features

This model supports the following features:

Tools Response Format Reasoning
Performance Summary

Qwen3.7 Flash, a vision-language reasoning model from Alibaba, demonstrates a balanced performance profile, particularly excelling in reliability and offering competitive pricing. Created on July 27, 2026, with a substantial context length of 1,000,000, it is designed for multimodal agents, visual coding, search, and computer interaction, leveraging strengths in object recognition and spatial understanding. The model exhibits exceptional reliability, achieving a perfect 100% success rate across benchmarks, indicating it consistently provides usable responses without technical failures. In terms of cost, Qwen3.7 Flash offers competitive solutions, ranking in the 63rd percentile. Its speed performance is moderate, placing it in the 25th percentile. Across the "Email Classification (Baseline)" benchmark, Qwen3.7 Flash achieved a high accuracy of 99.0%, ranking in the 86th percentile for this classification task. This indicates a strong ability to understand context, tone, and purpose in emails. While its cost for this task was $0.0062 (63rd percentile), its duration of 368416ms placed it in the 25th percentile for speed. Overall, Qwen3.7 Flash's key strengths lie in its high reliability and strong classification accuracy, making it a robust choice for applications requiring consistent and precise output, despite its moderate processing speed.

Model Pricing

Current Pricing

Feature Price (per 1M tokens)
Prompt $0.03
Completion $0.13
Input Cache Read $0.006
Input Cache Write $0.038

Price History

Available Endpoints
Provider Endpoint Name Context Length Pricing (Input) Pricing (Output)
Alibaba
Alibaba | qwen/qwen3.7-flash-20260727 1M $0.03 / 1M tokens $0.13 / 1M tokens
Benchmark Results
Benchmark Category Reasoning Strategy Free Executions Accuracy Cost Duration
Other Models by qwen