ModelPriceWatch.com
Last scan 2026-08-21 Models tracked 220 Providers 31 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

Gemini 3 Flash Preview

by Google

Preview flagship mid tier

Today's price · per 1M tokens

Input

$0.500

per 1M tokens

Output

$3.00

per 1M tokens

Blended

$1.13

blended $/1M — 3:1 weighted input:output

Cached input

$0.050

10% of input — prompt caching

Source: official Google pricing · read Jul 8, 2026 MODELPRICEWATCH.COM · 2026-08-21

Price receipt

We read Google's own pricing page on Jul 8, 2026 and found gemini-3-flash-preview listed with a price on the same row — that read is where the number above comes from, recorded as a new model. We keep the page text we read; its content hash is 33eeb06d22.

Overview

Preview release of Google's Gemini 3 Flash — a fast, low-cost multimodal model with a 1M-token context. Predecessor to Gemini 3.1/3.5 Flash. $0.50/$3 per 1M tokens.

Capabilities

struck through = not supported
Input 4/5
Text ✓ Image ✓ Audio ✓ Video ✓ PDF
Output 1/5
Text ✓ Image Audio Video Embedding
Features 4/9
Prompt caching ✓ Reasoning Coding Fast inference ✓ Long context ✓ Open weights Multimodal ✓ Web search Realtime

Benchmark performance

accuracy % · higher is better
Percentile vs all tracked models 53.8th2 independent measurements
Percentile per $/Mtok 48
GPQA Diamond vendor
90.4%
SWE-Bench Verified vendor
78%
Humanity's Last Exam vendor
33.7%

Every per-benchmark score we hold for this model is reported by its own vendor, so none is counted toward an accuracy average or any ranking. The percentile above is its standing across the independent composites below.

Independent composite scores
  • AA Intelligence Index: 27 (Artificial Analysis composite across reasoning, knowledge and coding evals)
  • LMSYS Chatbot Arena: 1472.8 ELO (Human preference)
How it stacks up
  • Ranks #62 of 161 comparably-measured models by percentile score, across 2 independent measurements
  • Ranks #50 of 161 comparably-tested models by normalized performance per dollar

167.8 tokens/sec output

Source: Artificial Analysis, Google model card, Google model card (no tools) · updated Jul 18, 2026 · See full rankings →

Specifications

Provider
Google
Context window
1M tokens
Modality
text, image, audio, video
Parameters
Proprietary
Open source
No — proprietary
Released
Dec 17, 2025
Status
Preview
Last updated
Jul 8, 2026
Tags
fastmultimodalpreview

Availability verified: Aug 9, 2026 — listed on Google's own page