inclusionAI: Ling 3.0 Flash VL (free)

Text input Image input Video input Text output
Author's Description

Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...

Key Specifications
Cost
$$
Context
131K
Parameters
124B (Rumoured)
Released
Sep 10, 2026
Speed
Ability
Reliability
Supported Parameters

This model supports the following parameters:

Temperature Include Reasoning Reasoning Logit Bias Presence Penalty Stop Tool Choice Min P Response Format Top P Frequency Penalty Tools Max Tokens Structured Outputs Seed
Features

This model supports the following features:

Reasoning Tools Response Format Structured Outputs
Performance Summary

inclusionAI: Ling 3.0 Flash VL demonstrates a balanced performance profile, particularly excelling in reliability and offering cost-effective solutions. Its speed performance is moderate, ranking in the 24th percentile across benchmarks. However, its exceptional reliability, with a 100% success rate, ensures consistent and usable responses. The model shows strong capabilities in Coding (95.0% accuracy, 90th percentile), Reasoning (97.5% accuracy, 80th percentile), and Mathematics (95.0% accuracy, 85th percentile), indicating robust analytical and problem-solving skills. Notably, it achieved perfect accuracy in Ethics, making it the most accurate model at its price point and among models of similar speed. While its General Knowledge (95.8% accuracy, 36th percentile) and Hallucinations (95.7% accuracy, 49th percentile) scores are respectable, they are not as standout as its other strengths. Instruction Following (69.1% accuracy, 64th percentile) is solid, and Email Classification (98.9% accuracy, 69th percentile) is strong. Overall, Ling 3.0 Flash VL is a highly reliable and cost-efficient model with significant strengths in coding, reasoning, and ethical decision-making.

Model Pricing

Current Pricing

Feature Price (per 1M tokens)
Prompt $0.06
Completion $0.18
Input Cache Read $0.012

Price History

Available Endpoints
Provider Endpoint Name Context Length Pricing (Input) Pricing (Output)
DeepInfra
DeepInfra | inclusionai/ling-3.0-flash-vl-20260910 131K $0.06 / 1M tokens $0.18 / 1M tokens
Benchmark Results
Benchmark Category Reasoning Strategy Free Executions Accuracy Cost Duration
Other Models by inclusionai