Google: Gemini 3.6 Flash

Image input File input Audio input Text input Video input Text output
Author's Description

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

Key Specifications
Cost
$$$$$
Context
1M
Released
Jul 21, 2026
Speed
Ability
Reliability
Supported Parameters

This model supports the following parameters:

Include Reasoning Max Tokens Structured Outputs Tool Choice Tools Response Format Seed Reasoning Top P Temperature
Features

This model supports the following features:

Response Format Tools Structured Outputs Reasoning
Performance Summary

Google's Gemini 3.6 Flash, created on July 21, 2026, is a high-efficiency model designed for coding, agentic workflows, and web/app development. It demonstrates competitive response times, ranking in the 54th percentile for speed across eight benchmarks. However, its pricing tends to be at premium levels, placing it in the 10th percentile for cost. A standout feature is its exceptional reliability, achieving a 100% success rate across all benchmarks, indicating minimal technical failures. The model exhibits perfect accuracy in Hallucinations, Reasoning, and Ethics benchmarks, often being the most accurate model at its price point and speed. It also shows strong performance in Coding (96.0% accuracy, 95th percentile) and Mathematics (95.0% accuracy, 84th percentile). While its Instruction Following accuracy is respectable at 83.0% (87th percentile), its General Knowledge (94.0% accuracy, 32nd percentile) and Email Classification (98.0% accuracy, 49th percentile) scores are more moderate. Its key strengths lie in its reliability and its ability to deliver perfect or near-perfect results in critical areas like hallucination prevention, complex reasoning, and ethical decision-making, making it suitable for tasks requiring high precision and trustworthiness.

Model Pricing

Current Pricing

Feature Price (per 1M tokens)
Prompt $1.5
Completion $7.5
Input Cache Read $0.15
Input Cache Write $0.0833
Internal Reasoning $7.5
Web Search $14000

Price History

Available Endpoints
Provider Endpoint Name Context Length Pricing (Input) Pricing (Output)
Google AI Studio
Google AI Studio | google/gemini-3.6-flash-20260721 1M $1.5 / 1M tokens $7.5 / 1M tokens
Google AI Studio
Google AI Studio | google/gemini-3.6-flash-20260721 1M $0.75 / 1M tokens $3.75 / 1M tokens
Google AI Studio
Google AI Studio | google/gemini-3.6-flash-20260721 1M $2.7 / 1M tokens $13.5 / 1M tokens
Google
Google | google/gemini-3.6-flash-20260721 1M $1.5 / 1M tokens $7.5 / 1M tokens
Google
Google | google/gemini-3.6-flash-20260721 1M $0.75 / 1M tokens $3.75 / 1M tokens
Google
Google | google/gemini-3.6-flash-20260721 1M $2.7 / 1M tokens $13.5 / 1M tokens
Benchmark Results
Benchmark Category Reasoning Strategy Free Executions Accuracy Cost Duration
Other Models by google