Qwen: Qwen3.8 Max Prime

Video input Image input Text input Text output
Author's Description

Qwen3.8 Max Prime is a higher-throughput variant of Qwen3.8 Max from Alibaba's Qwen team, served as a separate SKU at a higher price point. It accepts text, image, and video...

Key Specifications
Cost
$$$$$
Context
1M
Released
Sep 23, 2026
Speed
Ability
Reliability
Supported Parameters

This model supports the following parameters:

Max Tokens Frequency Penalty Temperature Tool Choice Include Reasoning Reasoning Top P Stop Seed Response Format Logprobs Tools Presence Penalty Structured Outputs Top Logprobs
Features

This model supports the following features:

Tools Response Format Structured Outputs Reasoning
Performance Summary

Qwen3.8 Max Prime, a higher-throughput variant from Alibaba's Qwen team, demonstrates moderate speed performance, ranking in the 36th percentile across benchmarks. Its pricing tends to be premium, positioned in the 10th percentile. A significant strength is its exceptional reliability, boasting a 98% success rate, indicating consistent and usable responses. The model exhibits strong performance in several key areas. It achieves perfect accuracy in General Knowledge, making it the most accurate model at its price point and among models of similar speed. Instruction Following and Mathematics also show high accuracy, at 88.1% (93rd percentile) and 95.5% (87th percentile) respectively, highlighting its capability in structured tasks and complex calculations. Coding performance is also robust at 94.9% accuracy. However, Qwen3.8 Max Prime struggles with Hallucinations, with an 86.5% accuracy (29th percentile), suggesting it occasionally fails to acknowledge uncertainty. Its Reasoning capabilities are also a notable weakness, with only 62.0% accuracy (34th percentile). Email Classification and Ethics benchmarks show average to good performance. Overall, Qwen3.8 Max Prime is a reliable model with strong knowledge and instruction-following capabilities, though its premium pricing and moderate speed are balanced by its exceptional reliability and specific areas of high accuracy.

Model Pricing

Current Pricing

Feature Price (per 1M tokens)
Prompt $4
Completion $12
Input Cache Read $0.5

Price History

Available Endpoints
Provider Endpoint Name Context Length Pricing (Input) Pricing (Output)
Alibaba
Alibaba | qwen/qwen3.8-max-prime-20260923 1M $4 / 1M tokens $12 / 1M tokens
Benchmark Results
Benchmark Category Reasoning Strategy Free Executions Accuracy Cost Duration
Other Models by qwen