ModelPriceWatch.com
Last scan 2026-09-15 Models tracked 259 Providers 32 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

Gemini 2.5 Pro

by Google

Deprecated mid tier mid tier
Deprecatedshuts down . Still purchasable until then, but the provider has announced end-of-life. Don't start new work on it. Migration target: Gemini 3.1 Pro. Full deprecation status & replacements →
Availability: Google publishes this as the EARLIEST possible retirement date.

Today's price · per 1M tokens

Input

$1.25

per 1M tokens

Output

$10.00

per 1M tokens

Blended

$3.44

blended $/1M — 3:1 weighted input:output

Source: official Google pricing · read MODELPRICEWATCH.COM · 2026-09-15

Price receipt

No confirmed receipt. We hold dated captures of Google's pricing page, but none of them tied Gemini 2.5 Pro to the price above on that page, so we do not claim this number is sourced. This model is deprecated and shuts down while it remains purchasable, confirm the current rate on the provider's page before relying on it.

Overview

Still available. 2M context window, strong reasoning. $1.25/$10 per 1M tokens.

Capabilities

struck through = not supported
Input 4/5
Text Image Audio Video PDF
Output 1/5
Text Image Audio Video Embedding
Features 3/9
Prompt caching Reasoning Coding Fast inference Long context Open weights Multimodal Web search Realtime

Benchmark performance

accuracy % · higher is better
Avg benchmark score61.2%
Perf per $/Mtok17.8
GPQA Diamond
86.4%
SWE-Bench Verified
50%
Terminal-Bench 2.1
32.6%
LiveCodeBench
69%
Humanity's Last Exam
21.6%
ARC-AGI 2
25%
AIME 2025
88%
MMMLU
89.2%
BFCL
72%
HumanEval
59.6%
MATH 500
80%
Independent composite scores
  • AA Intelligence Index: 16.7 (Artificial Analysis composite across reasoning, knowledge and coding evals)
  • AA-Omniscience Index: -16.4 (−100–100)
  • LMSYS Chatbot Arena: 1445.5 ELO (Human preference)
How it stacks up
  • Ranks #58 of 84 benchmarked models by average score
  • Ranks #86 of 181 comparably-measured models by percentile score, across 14 independent measurements
  • Ranks #145 of 181 comparably-tested models by normalized performance per dollar
  • Strongest at MMMLU — 89.2%, #5 of 34

191 tokens/sec output 30s latency to first token (TTFT)

Source: Vellum LLM Leaderboard · updated · See full rankings →

Specifications

Provider
Google
Context window
2M tokens
Modality
text, image, audio, video
Parameters
Proprietary
Open source
No — proprietary
Released
Status
Deprecated ends
Last updated
Tags
reasoningmid-tiermultimodal

Availability verified: per Google's own deprecation notice