Google: Gemini 3.8 Flash (batch)

Audio input Text input Image input Video input File input Text output
Author's Description

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

Key Specifications
Cost
$$$$
Context
1M
Released
Sep 02, 2026
Speed
Ability
Reliability
Supported Parameters

This model supports the following parameters:

Stop Tool Choice Response Format Include Reasoning Reasoning Tools Max Tokens Structured Outputs Seed
Features

This model supports the following features:

Reasoning Tools Response Format Structured Outputs
Performance Summary

Google's Gemini 3.8 Flash (batch), created on September 2, 2026, demonstrates a strong overall performance profile, particularly excelling in reliability and exhibiting significant gains in software engineering, agentic tasks, and multi-step reasoning. The model consistently delivers responses with a perfect 100% success rate across all 8 benchmarks, indicating exceptional reliability with minimal technical failures. In terms of speed, Gemini 3.8 Flash typically performs in the top tier, ranking in the 61st percentile. Its pricing is moderate, falling within the 37th percentile. Analyzing benchmark results, Gemini 3.8 Flash achieved perfect 100% accuracy in both Hallucinations (Baseline) and Ethics (Baseline), notably being the most accurate model at its price point and speed for these categories. It also shows high accuracy in Instruction Following (91.0%), General Knowledge (99.5%), Coding (96.0%), Reasoning (98.0%), and Mathematics (97.0%), consistently placing in high percentiles for accuracy. While its Email Classification accuracy (98.0%) is solid, it ranks in the 55th percentile, suggesting it's competitive but not a standout in this specific area. Key strengths include its robust reasoning capabilities, strong ethical adherence, and impressive performance in coding and mathematics, making it a versatile and dependable model for a wide range of applications. No significant weaknesses were identified across the evaluated benchmarks.

Model Pricing

Current Pricing

Feature Price (per 1M tokens)
Prompt $0.375
Completion $1.88
Input Cache Read $0.0375
Input Cache Write $0.0208
Internal Reasoning $1.88
Web Search $14000

Price History

Available Endpoints
Provider Endpoint Name Context Length Pricing (Input) Pricing (Output)
Google
Google | google/gemini-3.8-flash-20260902 1M $0.375 / 1M tokens $1.88 / 1M tokens
Google AI Studio
Google AI Studio | google/gemini-3.8-flash-20260902 1M $0.375 / 1M tokens $1.88 / 1M tokens
Google
Google | google/gemini-3.8-flash-20260902 1M $0.75 / 1M tokens $3.75 / 1M tokens
Google AI Studio
Google AI Studio | google/gemini-3.8-flash-20260902 1M $0.75 / 1M tokens $3.75 / 1M tokens
Google AI Studio
Google AI Studio | google/gemini-3.8-flash-20260902 1M $1.35 / 1M tokens $6.75 / 1M tokens
Google
Google | google/gemini-3.8-flash-20260902 1M $1.35 / 1M tokens $6.75 / 1M tokens
Benchmark Results
Benchmark Category Reasoning Strategy Free Executions Accuracy Cost Duration
Other Models by google