ModelPriceWatch.com
Last scan 2026-08-21 Models tracked 220 Providers 31 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

Gemini 3 Flash Preview vs Command A

Side-by-side comparison of API pricing, specs, benchmarks, and capabilities

Gemini 3 Flash Preview is 74.3% cheaper on blended cost ($1.13 vs $4.38/Mtok)
Specification
Overview
StatusPreview flagship Current flagship
Released Dec 17, 2025 Mar 13, 2025
Pricing per million tokens
Input $0.500/Mtok $2.50/Mtok
Output $3.00/Mtok $10.00/Mtok
Blended avg $1.13/Mtok $4.38/Mtok
Cached input $0.050/Mtok
Specifications
Context window 1M tokens 256K tokens
Parameters Proprietary 111B
Speed (TPS)
Modalities
Input
textimageaudiovideo
text
Benchmarks sources: Artificial Analysis, Google model card, Google model card (no tools) / Cohere model card
GPQA Diamond 90.4 vendor 50 vendor
SWE-Bench Verified 78 vendor 25 vendor
Humanity's Last Exam 33.7 vendor 15 vendor
ARC-AGI 2 12 vendor
AIME 2025 45 vendor
MMMLU 74 vendor
BFCL 60 vendor
HumanEval 78 vendor
MATH 500 62 vendor
Independent composite scores each on its own scale — not part of the average above
LMSYS Chatbot Arena 1472.8 ELO 1353.8 ELO
AA Intelligence Index 27
Providers
Available from
Google — $0.500/$3.00/Mtok
Cohere — $2.50/$10.00/Mtok

Cost at scale

1M tokens · 50/50 input/output
Projected cost of Gemini 3 Flash Preview vs Command A at increasing token volumes
VolumeGemini 3 Flash PreviewCommand ASavings
1M tokens $1.13 $4.38 $3.25 (74.3%)
10M tokens $11.25 $43.75 $32.5 (74.3%)
100M tokens $112.5 $437.5 $325 (74.3%)
1000M tokens $1125 $4375 $3250 (74.3%)

When to pick which

distilled from the pricing and spec data above
Pick Gemini 3 Flash Preview if…
  • cost dominates: $1.13/Mtok blended vs $4.38 — 74.3% less on the same 50/50 token mix
  • your workload re-reads context (agents, RAG, long chats): cached input costs $0.050/Mtok — 90% off its list input price, a discount Command A doesn't offer
  • you need the longer context: 1M tokens vs 256K (4.1×)
Pick Command A if…
  • this pairing gives Command A no edge on price, context, cache discount, or measured performance — pick it only for qualitative fit (ecosystem, compliance, model behavior on your prompts)
Try Gemini 3 Flash Preview on Google

The cheaper option here — Gemini 3 Flash Preview costs $1.13/Mtok blended on Google.

Get API key →

Summary

Gemini 3 Flash Preview by Google costs $0.500/Mtok input and $3.00/Mtok output, with a 1M-token context window. It supports text, image, audio, video input.

Command A by Cohere costs $2.50/Mtok input and $10.00/Mtok output, with a 256K-token context window. It supports text input.

On a blended cost basis, Gemini 3 Flash Preview is 74.3% cheaper than Command A.It also has a larger context window.

Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.

More comparisons

pairs sharing a model with this page