ModelPriceWatch.com
Last scan 2026-08-01 Models tracked 181 Providers 27 Cheapest paid Granite 4.0 Micro $0.017/Mtok in Every price links to its source

Gemini 3.1 Flash-Lite

by Google

Current budget cheap tier

Today's price · per 1M tokens

Input

$0.250

per 1M tokens

Output

$1.50

per 1M tokens

Blended

$0.875

avg of input & output $/1M

Source: official Google pricing · read Jun 25, 2026 MODELPRICEWATCH.COM · 2026-08-01

Overview

Ultra-cheap, fast model. $0.10/$0.40 per 1M tokens with full multimodal support.

Capabilities

struck through = not supported
Input 4/5
Text ✓ Image ✓ Audio ✓ Video ✓ PDF
Output 1/5
Text ✓ Image Audio Video Embedding
Features 3/9
Prompt caching Reasoning Coding Fast inference ✓ Long context ✓ Open weights Multimodal ✓ Web search Realtime

Benchmark performance

accuracy % · higher is better
GPQA Diamond vendor
50%
SWE-Bench Verified vendor
30%
Humanity's Last Exam vendor
15%
ARC-AGI 2 vendor
18%
AIME 2025 vendor
50%
MMMLU vendor
72%
BFCL vendor
60%
HumanEval vendor
78%
MATH 500 vendor
65%

Every per-benchmark score we hold for this model is reported by its own vendor, so none is counted toward an accuracy average or any ranking.

600 tokens/sec output

Source: Google model card · updated Apr 15, 2026 · See full rankings →

Specifications

Provider
Google
Context window
1M tokens
Modality
text, image, audio, video
Parameters
Proprietary
Open source
No — proprietary
Released
Jan 1, 2026
Status
Current
Last updated
Jun 25, 2026
Tags
fastmultimodalbudget

Availability verified: Not re-verified. We have no recent evidence that Google still lists this model — treat availability as unconfirmed.