Gemini 3.6 Flash
by Google
Today's price · per 1M tokens
Input
$0.750
per 1M tokens
Output
$3.75
per 1M tokens
Blended
$1.50
blended $/1M — 3:1 weighted input:output
Cached input
$0.075
10% of input — prompt caching
How this price is scoped: Google's Gemini API pricing page prints a dated, two-phase schedule for this model: $0.75 per 1M input tokens through December 31, 2026, rising to $1.50 on January 1, 2027, and $3.75 per 1M output tokens through December 31, 2026, rising to $7.50. Context caching follows the same dates ($0.075, then $0.15). We publish the rate you would pay today and re-verify it before the changeover date. Checked 2026-08-13 against ai.google.dev/gemini-api/docs/pricing; the same page listed a flat $1.50/$7.50 on 2026-08-09, so this is a genuine 50% price cut rather than a correction to an earlier mistake.
Price receipt
We read Google's own pricing page on and found Gemini 3.6 Flash listed with a price on the same card — that read is where the number above comes from, recorded as a price drop. We keep the page text we read; its content hash is 08097b2680.
Overview
Google's fast multimodal flagship line, successor to Gemini 3.5 Flash and now joined by Gemini 3.7 Flash. Currently $0.75/1M input and $3.75/1M output on Google's dated schedule, reverting to $1.50/$7.50 on 2027-01-01. 1M-token context with native image, audio and video input.
Capabilities
struck through = not supportedBenchmark performance
accuracy % · higher is betterEvery per-benchmark score we hold for this model is reported by its own vendor, so none is counted toward an accuracy average or any ranking. The percentile above is its standing across the independent composites below.
- AA Intelligence Index: 51.6 (Artificial Analysis composite across reasoning, knowledge and coding evals)
- AA Agentic Index: 40.5 (Tool use, planning, autonomy)
- AA-Omniscience Index: 22.1 (−100–100)
- GDPval-AA v2: 1422.3 ELO (Real-world work tasks, human baseline = 1000)
- LMSYS Chatbot Arena: 1480.2 ELO (Human preference)
- Ranks #33 of 176 comparably-measured models by percentile score, across 5 independent measurements
- Ranks #55 of 176 comparably-tested models by normalized performance per dollar
189.8 tokens/sec output
Source: Artificial Analysis, Google DeepMind Gemini 3.6 Flash model card, LMSYS Chatbot Arena (UC Berkeley) · updated · See full rankings →
Specifications
- Provider
- Context window
- 1M tokens
- Modality
- text, image, audio, video
- Parameters
- Proprietary
- Open source
- No — proprietary
- Released
- Status
- Current
- Last updated
- Tags
Availability verified: — listed on Google's own page