Author's Description
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
Key Specifications
Supported Parameters
This model supports the following parameters:
Features
This model supports the following features:
Performance Summary
Z.ai's GLM 5.3 Flash, created on August 26, 2026, is a native multimodal model designed for efficient coding and long-horizon agent tasks, leveraging a hybrid sparse and linear attention architecture for accurate long-context behavior. The model demonstrates moderate speed performance, ranking in the 34th percentile across seven benchmarks. It generally offers cost-effective solutions, placing in the 69th percentile for price. A standout feature is its exceptional reliability, boasting a 99% success rate, indicating consistent and usable responses. In terms of benchmark performance, GLM 5.3 Flash exhibits perfect accuracy in Hallucinations (100.0%) and General Knowledge (100.0%), positioning it as a top performer in these areas, often being the most accurate model at its price point and speed. It also shows strong capabilities in Reasoning (98.0% accuracy, 91st percentile) and Coding (95.0% accuracy, 90th percentile), making it well-suited for its intended applications. While its Instruction Following is solid at 77.0% accuracy (79th percentile), its Ethics performance is comparatively lower at 96.0% accuracy (25th percentile), suggesting a potential area for improvement despite still being a high score. Email Classification is strong at 98.0% accuracy. Overall, its key strengths lie in its accuracy for knowledge-based and reasoning tasks, coupled with high reliability and cost-effectiveness.
Model Pricing
Current Pricing
| Feature | Price (per 1M tokens) |
|---|---|
| Prompt | $0.075 |
| Completion | $0.25 |
| Input Cache Read | $0.015 |
Price History
Available Endpoints
| Provider | Endpoint Name | Context Length | Pricing (Input) | Pricing (Output) |
|---|---|---|---|---|
|
Z.AI
|
Z.AI | z-ai/glm-5.3-flash-20260826 | 1M | $0.075 / 1M tokens | $0.25 / 1M tokens |
Benchmark Results
| Benchmark | Category | Reasoning | Strategy | Free | Executions | Accuracy | Cost | Duration |
|---|
Other Models by z-ai
|
|
Released | Params | Context |
|
Speed | Ability | Cost |
|---|---|---|---|---|---|---|---|
| Z.ai: GLM 5.3 | Aug 18, 2026 | ~5.3B | 1M |
Text input
Text output
|
★★ | ★★★★★ | $$$$$ |
| Z.ai: GLM 5.2 (free) | Jun 16, 2026 | ~5.2B | 1M |
Text input
Text output
|
★ | ★★★★ | $$$$$ |
| Z.ai: GLM 5.1 | Apr 07, 2026 | — | 202K |
Text input
Text output
|
★ | ★★★★★ | $$$$$ |
| Z.ai: GLM 5V Turbo | Apr 01, 2026 | — | 202K |
Text input
Image input
Video input
Text output
|
★★ | ★★★★ | $$$$$ |
| Z.ai: GLM 5 Turbo Unavailable | Mar 15, 2026 | — | 202K |
Text input
Text output
|
— | — | $$$$$ |
| Z.ai: GLM 5 Turbo | Mar 15, 2026 | — | 202K |
Text input
Text output
|
★★ | ★★★★★ | $$$$$ |
| Z.ai: GLM 5 | Feb 11, 2026 | — | 202K |
Text input
Text output
|
★ | ★★★★ | $$$$$ |
| Z.ai: GLM 5 Unavailable | Feb 11, 2026 | — | 204K |
Text input
Text output
|
★ | ★★★★ | $ |
| Z.ai: GLM 4.7 Flash | Jan 19, 2026 | ~30B | 202K |
Text input
Text output
|
★ | ★★★ | $$$$ |
| Z.ai: GLM 4.7 | Dec 21, 2025 | — | 202K |
Text input
Text output
|
★ | ★★★★ | $$$$$ |
| Z.ai: GLM 4.6V | Dec 08, 2025 | — | 131K |
Text input
Image input
Video input
Text output
|
★★ | ★★★★★ | $$$$ |
| Z.ai: GLM 4.6 | Sep 30, 2025 | — | 202K |
Text input
Text output
|
★ | ★★★ | $$$$$ |
| Z.ai: GLM 4.6 (exacto) Unavailable | Sep 30, 2025 | — | 202K |
Text input
Text output
|
— | — | $$$$ |
| Z.ai: GLM 4.5V | Aug 11, 2025 | ~106B | 65K |
Text input
Image input
Text output
|
★★ | ★★★ | $$$$ |
| Z.ai: GLM 4.5 | Jul 25, 2025 | — | 131K |
Text input
Text output
|
★ | ★★★★ | $$$$$ |
| Z.ai: GLM 4.5 Air | Jul 25, 2025 | — | 131K |
Text input
Text output
|
★ | ★★ | $$$$ |
| Z.ai: GLM 4 32B Unavailable | Jul 24, 2025 | 32B | 128K |
Text input
Text output
|
★★★ | ★ | $$ |