Author's Description
Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic planning, and reports strong results across knowledge, reasoning, coding, alignment, and multilingual evaluations. Compared with prior Qwen3 variants, it emphasizes stability under long chains of thought and efficient scaling during inference, and it is tuned to follow complex instructions while reducing repetitive or off-task behavior. The model is suitable for agent frameworks and tool use (function calling), retrieval-heavy workflows, and standardized benchmarking where step-by-step solutions are required. It supports long, detailed completions and leverages throughput-oriented techniques (e.g., multi-token prediction) for faster generation. Note that it operates in thinking-only mode.
Key Specifications
Supported Parameters
This model supports the following parameters:
Features
This model supports the following features:
Performance Summary
Qwen3-Next-80B-A3B-Thinking is a reasoning-focused chat model designed for complex multi-step problems. Its speed performance tends to be slower, ranking in the 15th percentile across benchmarks, indicating longer response times. In terms of cost, it is positioned at premium pricing levels, falling into the 5th percentile. However, the model demonstrates exceptional reliability with a 99% success rate, consistently providing usable responses. Across benchmarks, the model exhibits strong performance in Coding (94.9% accuracy, 90th percentile), Reasoning (96.0% accuracy, 81st percentile), and Ethics (100% accuracy, achieving perfect scores and being the most accurate at its price point and speed). General Knowledge and Email Classification also show high accuracy at 99.5% and 99.0% respectively. A notable strength is its low hallucination rate (98.0% accuracy). The primary weakness lies in Instruction Following, where it achieved only 14.7% accuracy, placing it in the 21st percentile. Mathematics performance is moderate at 88.9% accuracy. Its "thinking-only" mode is a key feature for structured problem-solving.
Model Pricing
Current Pricing
| Feature | Price (per 1M tokens) |
|---|---|
| Prompt | $0.0975 |
| Completion | $0.78 |
Price History
Available Endpoints
| Provider | Endpoint Name | Context Length | Pricing (Input) | Pricing (Output) |
|---|---|---|---|---|
|
Alibaba
|
Alibaba | qwen/qwen3-next-80b-a3b-thinking-2509 | 131K | $0.0975 / 1M tokens | $0.78 / 1M tokens |
|
Novita
|
Novita | qwen/qwen3-next-80b-a3b-thinking-2509 | 131K | $0.0975 / 1M tokens | $0.78 / 1M tokens |
|
Chutes
|
Chutes | qwen/qwen3-next-80b-a3b-thinking-2509 | 262K | $0.0975 / 1M tokens | $0.78 / 1M tokens |
|
DeepInfra
|
DeepInfra | qwen/qwen3-next-80b-a3b-thinking-2509 | 262K | $0.0975 / 1M tokens | $0.78 / 1M tokens |
|
Hyperbolic
|
Hyperbolic | qwen/qwen3-next-80b-a3b-thinking-2509 | 262K | $0.0975 / 1M tokens | $0.78 / 1M tokens |
|
GMICloud
|
GMICloud | qwen/qwen3-next-80b-a3b-thinking-2509 | 262K | $0.0975 / 1M tokens | $0.78 / 1M tokens |
|
AtlasCloud
|
AtlasCloud | qwen/qwen3-next-80b-a3b-thinking-2509 | 262K | $0.15 / 1M tokens | $1.5 / 1M tokens |
|
NCompass
|
NCompass | qwen/qwen3-next-80b-a3b-thinking-2509 | 262K | $0.0975 / 1M tokens | $0.78 / 1M tokens |
|
Together
|
Together | qwen/qwen3-next-80b-a3b-thinking-2509 | 262K | $0.0975 / 1M tokens | $0.78 / 1M tokens |
|
Parasail
|
Parasail | qwen/qwen3-next-80b-a3b-thinking-2509 | 262K | $0.0975 / 1M tokens | $0.78 / 1M tokens |
|
Google
|
Google | qwen/qwen3-next-80b-a3b-thinking-2509 | 262K | $0.15 / 1M tokens | $1.2 / 1M tokens |
|
Parasail
|
Parasail | qwen/qwen3-next-80b-a3b-thinking-2509 | 262K | $0.0975 / 1M tokens | $0.78 / 1M tokens |
|
Chutes
|
Chutes | qwen/qwen3-next-80b-a3b-thinking-2509 | 262K | $0.0975 / 1M tokens | $0.78 / 1M tokens |
|
Novita
|
Novita | qwen/qwen3-next-80b-a3b-thinking-2509 | 131K | $0.15 / 1M tokens | $1.5 / 1M tokens |
|
Nebius
|
Nebius | qwen/qwen3-next-80b-a3b-thinking-2509 | 128K | $0.15 / 1M tokens | $1.2 / 1M tokens |
Benchmark Results
| Benchmark | Category | Reasoning | Strategy | Free | Executions | Accuracy | Cost | Duration |
|---|
Other Models by qwen
|
|
Released | Params | Context |
|
Speed | Ability | Cost |
|---|---|---|---|---|---|---|---|
| Qwen: Qwen3.5-9B | Mar 10, 2026 | 9B | 262K |
Text input
Video input
Image input
Text output
|
★ | ★★★ | $$$$ |
| Qwen: Qwen3.5-35B-A3B | Feb 25, 2026 | 35B | 262K |
Text input
Video input
Image input
Text output
|
★ | ★ | $$$$$ |
| Qwen: Qwen3.5-27B | Feb 25, 2026 | 27B | 262K |
Text input
Video input
Image input
Text output
|
★ | ★ | $$$$$ |
| Qwen: Qwen3.5-122B-A10B | Feb 25, 2026 | 122B | 262K |
Text input
Video input
Image input
Text output
|
★ | ★ | $$$$$ |
| Qwen: Qwen3.5-Flash | Feb 25, 2026 | — | 1M |
Text input
Video input
Image input
Text output
|
★★ | ★ | $$$$ |
| Qwen: Qwen3.5 Plus 2026-02-15 | Feb 16, 2026 | — | 1M |
Text input
Video input
Image input
Text output
|
★★★★ | ★ | $$$ |
| Qwen: Qwen3.5 397B A17B | Feb 15, 2026 | 397B | 262K |
Text input
Video input
Image input
Text output
|
★ | ★★★★★ | $$$$$ |
| Qwen: Qwen3 Max Thinking | Feb 09, 2026 | — | 262K |
Text input
Text output
|
★★ | ★★★★ | $$$$$ |
| Qwen: Qwen3 Coder Next | Feb 03, 2026 | ~80B | 262K |
Text input
Text output
|
★ | ★★★★ | $$$$ |
| Qwen: Qwen3 VL 32B Instruct | Oct 23, 2025 | 32B | 262K |
Text input
Image input
Text output
|
★★★ | ★★★★★ | $$ |
| Qwen: Qwen3 VL 8B Thinking | Oct 14, 2025 | 8B | 131K |
Text input
Image input
Text output
|
★ | ★ | $$$$$ |
| Qwen: Qwen3 VL 8B Instruct | Oct 14, 2025 | 8B | 131K |
Text input
Image input
Text output
|
★ | ★★ | $$$ |
| Qwen: Qwen3 VL 30B A3B Thinking | Oct 06, 2025 | 30B | 262K |
Text input
Image input
Text output
|
★ | ★★★ | $$$$ |
| Qwen: Qwen3 VL 30B A3B Instruct | Oct 06, 2025 | 30B | 131K |
Text input
Image input
Text output
|
— | — | $$$ |
| Qwen: Qwen3 VL 235B A22B Thinking | Sep 23, 2025 | 235B | 131K |
Text input
Image input
Text output
|
★ | ★ | $$$$$ |
| Qwen: Qwen3 VL 235B A22B Instruct | Sep 23, 2025 | 235B | 131K |
Text input
Image input
Text output
|
★★★★ | ★★★★ | $$$ |
| Qwen: Qwen3 Max | Sep 23, 2025 | — | 262K |
Text input
Text output
|
★★★★ | ★★★★ | $$$$ |
| Qwen: Qwen3 Coder Plus | Sep 23, 2025 | ~480B | 1M |
Text input
Text output
|
★★★★ | ★★★★ | $$$$ |
| Qwen: Qwen3 Coder Flash | Sep 17, 2025 | — | 1M |
Text input
Text output
|
★★★★ | ★★★ | $$$ |
| Qwen: Qwen3 Next 80B A3B Instruct | Sep 11, 2025 | 80B | 262K |
Text input
Text output
|
★★★★ | ★★★★ | $$$$ |
| Qwen: Qwen Plus 0728 | Sep 08, 2025 | ~20B | 1M |
Text input
Text output
|
★★★★★ | ★★★ | $$ |
| Qwen: Qwen3 30B A3B Thinking 2507 | Aug 28, 2025 | 30B | 262K |
Text input
Text output
|
★★ | ★★★ | $$$ |
| Qwen: Qwen3 Coder 30B A3B Instruct | Jul 31, 2025 | 30B | 262K |
Text input
Text output
|
★★★★ | ★★★ | $$ |
| Qwen: Qwen3 30B A3B Instruct 2507 | Jul 29, 2025 | 30B | 131K |
Text input
Text output
|
★★★★ | ★★★ | $$$ |
| Qwen: Qwen3 235B A22B Thinking 2507 | Jul 25, 2025 | 235B | 131K |
Text input
Text output
|
★ | ★★★★ | $$$$$ |
| Qwen: Qwen3 Coder 480B A35B | Jul 22, 2025 | 480B | 1M |
Text input
Text output
|
★★ | ★★★ | $$$ |
| Qwen: Qwen3 Coder 480B A35B (exacto) Unavailable | Jul 22, 2025 | 480B | 262K |
Text input
Text output
|
— | — | $$$$ |
| Qwen: Qwen3 235B A22B Instruct 2507 | Jul 21, 2025 | 235B | 262K |
Text input
Text output
|
★★ | ★★★ | $$$ |
| Qwen: Qwen3 4B Unavailable | Apr 30, 2025 | 4B | 131K |
Text input
Text output
|
★ | ★★★ | $$$ |
| Qwen: Qwen3 30B A3B | Apr 28, 2025 | 30B | 16K |
Text input
Text output
|
★★ | ★★★★ | $$$ |
| Qwen: Qwen3 8B | Apr 28, 2025 | 8B | 128K |
Text input
Text output
|
★ | ★★★ | $$$ |
| Qwen: Qwen3 14B | Apr 28, 2025 | 14B | 40K |
Text input
Text output
|
★★ | ★★★ | $$$ |
| Qwen: Qwen3 32B | Apr 28, 2025 | 32B | 40K |
Text input
Text output
|
★ | ★★★★ | $$$ |
| Qwen: Qwen3 235B A22B | Apr 28, 2025 | 235B | 40K |
Text input
Text output
|
★ | ★★★★ | $$$$ |
| Qwen: Qwen2.5 Coder 7B Instruct | Apr 15, 2025 | 7B | 32K |
Text input
Text output
|
— | — | $ |
| Qwen: Qwen2.5 VL 32B Instruct | Mar 24, 2025 | 32B | 128K |
Text input
Image input
Text output
|
★ | ★★★ | $$$ |
| Qwen: QwQ 32B | Mar 05, 2025 | 32B | 131K |
Text input
Text output
|
★ | ★★ | $$$ |
| Qwen: Qwen VL Plus | Feb 04, 2025 | — | 131K |
Text input
Image input
Text output
|
★★★★ | ★★ | $$$ |
| Qwen: Qwen VL Max | Feb 01, 2025 | — | 131K |
Text input
Image input
Text output
|
★★★ | ★★★ | $$$$ |
| Qwen: Qwen-Turbo | Feb 01, 2025 | — | 131K |
Text input
Text output
|
★★★★★ | ★★★★ | $$ |
| Qwen: Qwen2.5 VL 72B Instruct | Feb 01, 2025 | 72B | 128K |
Text input
Image input
Text output
|
★★★★ | ★★★★ | $$ |
| Qwen: Qwen-Plus | Feb 01, 2025 | — | 1M |
Text input
Text output
|
★★★★ | ★★★★ | $$$ |
| Qwen: Qwen-Max | Feb 01, 2025 | — | 32K |
Text input
Text output
|
★★★★ | ★★★★ | $$$$ |
| Qwen: QwQ 32B Preview Unavailable | Nov 27, 2024 | 32B | 32K |
Text input
Text output
|
— | ★ | $$ |
| Qwen2.5 Coder 32B Instruct | Nov 11, 2024 | ~500B | 32K |
Text input
Text output
|
★★★★★ | ★★★★★ | $ |
| Qwen: Qwen2.5 7B Instruct | Oct 15, 2024 | ~500B | 32K |
Text input
Text output
|
★ | ★★ | $ |
| Qwen2.5 72B Instruct | Sep 18, 2024 | ~500B | 32K |
Text input
Text output
|
★★★ | ★★ | $$ |
| Qwen: Qwen2.5-VL 7B Instruct Unavailable | Aug 27, 2024 | ~500B | 32K |
Text input
Image input
Text output
|
★★★★ | ★★ | $$ |
| Qwen 2 72B Instruct Unavailable | Jun 06, 2024 | ~500B | 32K |
Text input
Text output
|
★★★★ | ★★ | $$$$ |