inclusionAI: Ling 3.1 Flash

Text input Text output
Author's Description

Ling 3.1 Flash is a hybrid reasoning mixture-of-experts model from inclusionAI, with 25B active parameters out of 560B total.

Key Specifications
Cost
$
Context
262K
Parameters
560B (Rumoured)
Released
Oct 02, 2026
Speed
★
Ability
★★★
Reliability
★★
Supported Parameters

This model supports the following parameters:

Top P Seed Logprobs Stop Frequency Penalty Tools Temperature Tool Choice Reasoning Include Reasoning Presence Penalty Top Logprobs Max Tokens
Features

This model supports the following features:

Tools Reasoning
Performance Summary

inclusionAI's Ling 3.1 Flash, a hybrid reasoning mixture-of-experts model with 25B active parameters, demonstrates strong overall performance, particularly in reliability. The model exhibits strong reliability with a 93% success rate across benchmarks, indicating consistent and usable responses. In terms of speed, Ling 3.1 Flash shows moderate performance, ranking in the 31st percentile. Price data is currently unavailable, suggesting potential free-tier usage. Analyzing benchmark results, Ling 3.1 Flash excels in Instruction Following, achieving 84.0% accuracy and ranking in the 88th percentile, notably placing in the top 3 for cost efficiency in this category. Its Coding capabilities are also impressive, with 95.0% accuracy (89th percentile). While its Reasoning accuracy is 78.0% (50th percentile), its Mathematics performance is solid at 91.9% accuracy (51st percentile). A notable weakness is its relatively longer duration in Coding and Mathematics benchmarks, placing in the 22nd and 21st percentiles for speed, respectively. Overall, Ling 3.1 Flash is a highly reliable model with strong instruction following and coding abilities, making it a robust choice for tasks requiring precision and complex code generation.

Model Pricing

Current Pricing

Feature Price (per 1M tokens)
Prompt $0
Completion $0

Price History

Available Endpoints
Provider Endpoint Name Context Length Pricing (Input) Pricing (Output)
Novita
Novita | inclusionai/ling-3.1-flash-20261002 262K $0 / 1M tokens $0 / 1M tokens
Benchmark Results
Benchmark Category Reasoning Strategy Free Executions Accuracy Cost Duration
Other Models by inclusionai