Google: Gemini 3.5 Flash-Lite

Image input File input Audio input Text input Video input Text output
Author's Description

Gemini 3.5 Flash-Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

Key Specifications
Cost
$$$
Context
1M
Released
Jul 21, 2026
Speed
Ability
Reliability
Supported Parameters

This model supports the following parameters:

Include Reasoning Max Tokens Structured Outputs Tool Choice Tools Response Format Seed Reasoning Top P Temperature
Features

This model supports the following features:

Response Format Tools Structured Outputs Reasoning
Performance Summary

Google's Gemini 3.5 Flash-Lite, a high-efficiency model designed for focused subagent tasks, demonstrates exceptional performance across various metrics. It consistently ranks among the fastest models, achieving the 94th percentile across seven benchmarks, making it highly suitable for time-sensitive applications. The model also offers competitive pricing, positioned at the 55th percentile, balancing cost-effectiveness with its robust capabilities. Notably, Gemini 3.5 Flash-Lite exhibits outstanding reliability, boasting a 100% success rate across all benchmarks, indicating minimal technical failures and consistent responsiveness. In terms of specific benchmark performance, the model shows remarkable accuracy in General Knowledge and Ethics, achieving perfect 100% scores and being recognized as the most accurate at its price point and speed for these categories. Its Coding performance is also strong at 93% accuracy, making it the most accurate among models of similar speed. While its Hallucinations accuracy is a respectable 98%, its Instruction Following accuracy of 70% suggests an area for potential improvement, particularly in complex, multi-layered directives. Email Classification, at 97% accuracy, is solid, and Reasoning at 82% is adequate. Overall, Gemini 3.5 Flash-Lite's key strengths lie in its speed, reliability, and high accuracy in knowledge-based and ethical reasoning tasks, making it a powerful tool for agentic workflows.

Model Pricing

Current Pricing

Feature Price (per 1M tokens)
Prompt $0.3
Completion $2.5
Input Cache Read $0.03
Input Cache Write $0.0833
Internal Reasoning $2.5
Web Search $14000

Price History

Available Endpoints
Provider Endpoint Name Context Length Pricing (Input) Pricing (Output)
Google AI Studio
Google AI Studio | google/gemini-3.5-flash-lite-20260721 1M $0.3 / 1M tokens $2.5 / 1M tokens
Google AI Studio
Google AI Studio | google/gemini-3.5-flash-lite-20260721 1M $0.15 / 1M tokens $1.25 / 1M tokens
Google AI Studio
Google AI Studio | google/gemini-3.5-flash-lite-20260721 1M $0.54 / 1M tokens $4.5 / 1M tokens
Google
Google | google/gemini-3.5-flash-lite-20260721 1M $0.3 / 1M tokens $2.5 / 1M tokens
Google
Google | google/gemini-3.5-flash-lite-20260721 1M $0.15 / 1M tokens $1.25 / 1M tokens
Google
Google | google/gemini-3.5-flash-lite-20260721 1M $0.54 / 1M tokens $4.5 / 1M tokens
Benchmark Results
Benchmark Category Reasoning Strategy Free Executions Accuracy Cost Duration
Other Models by google