ModelPriceWatch.com
Last scan 2026-09-04 Models tracked 244 Providers 31 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

Gemini 3.6 Flash

by Google

Current new · 45d Intro price flagship mid tier

Today's price · per 1M tokens

Input

$0.750

per 1M tokens

Output

$3.75

per 1M tokens

Blended

$1.50

blended $/1M — 3:1 weighted input:output

Cached input

$0.075

10% of input — prompt caching

How this price is scoped: Google's Gemini API pricing page prints a dated, two-phase schedule for this model: $0.75 per 1M input tokens through December 31, 2026, rising to $1.50 on January 1, 2027, and $3.75 per 1M output tokens through December 31, 2026, rising to $7.50. Context caching follows the same dates ($0.075, then $0.15). We publish the rate you would pay today and re-verify it before the changeover date. Checked 2026-08-13 against ai.google.dev/gemini-api/docs/pricing; the same page listed a flat $1.50/$7.50 on 2026-08-09, so this is a genuine 50% price cut rather than a correction to an earlier mistake.

Source: official Google pricing · read MODELPRICEWATCH.COM · 2026-09-04

Price receipt

We read Google's own pricing page on and found Gemini 3.6 Flash listed with a price on the same card — that read is where the number above comes from, recorded as a price drop. We keep the page text we read; its content hash is 08097b2680.

Overview

Google's fast multimodal flagship line, successor to Gemini 3.5 Flash and now joined by Gemini 3.7 Flash. Currently $0.75/1M input and $3.75/1M output on Google's dated schedule, reverting to $1.50/$7.50 on 2027-01-01. 1M-token context with native image, audio and video input.

Capabilities

struck through = not supported
Input 4/5
Text Image Audio Video PDF
Output 1/5
Text Image Audio Video Embedding
Features 4/9
Prompt caching Reasoning Coding Fast inference Long context Open weights Multimodal Web search Realtime

Benchmark performance

accuracy % · higher is better
Percentile vs all tracked models 69.3th5 independent measurements
Percentile per $/Mtok 46.2
Terminal-Bench 2.1 vendor
78%
OSWorld-Verified vendor
83%

Every per-benchmark score we hold for this model is reported by its own vendor, so none is counted toward an accuracy average or any ranking. The percentile above is its standing across the independent composites below.

Independent composite scores
  • AA Intelligence Index: 51.6 (Artificial Analysis composite across reasoning, knowledge and coding evals)
  • AA Agentic Index: 40.5 (Tool use, planning, autonomy)
  • AA-Omniscience Index: 22.1 (−100–100)
  • GDPval-AA v2: 1422.3 ELO (Real-world work tasks, human baseline = 1000)
  • LMSYS Chatbot Arena: 1480.2 ELO (Human preference)
How it stacks up
  • Ranks #33 of 176 comparably-measured models by percentile score, across 5 independent measurements
  • Ranks #55 of 176 comparably-tested models by normalized performance per dollar

189.8 tokens/sec output

Source: Artificial Analysis, Google DeepMind Gemini 3.6 Flash model card, LMSYS Chatbot Arena (UC Berkeley) · updated · See full rankings →

Specifications

Provider
Google
Context window
1M tokens
Modality
text, image, audio, video
Parameters
Proprietary
Open source
No — proprietary
Released
Status
Current
Last updated
Tags
fastmultimodalflagship

Availability verified: listed on Google's own page