ModelPriceWatch.com
Last scan 2026-08-12 Models tracked 203 Providers 31 Cheapest paid Granite 4.0 Micro $0.017/Mtok in Every price links to its source

Command A vs Gemini 3.5 Flash

Side-by-side comparison of API pricing, specs, benchmarks, and capabilities

Gemini 3.5 Flash is 22.9% cheaper on blended cost ($3.38 vs $4.38/Mtok)
Specification
Overview
StatusCurrent flagship Current flagship
Released Mar 13, 2025 May 19, 2026
Pricing per million tokens
Input $2.50/Mtok $1.50/Mtok
Output $10.00/Mtok $9.00/Mtok
Blended avg $4.38/Mtok $3.38/Mtok
Specifications
Context window 256K tokens 1M tokens
Parameters 111B Proprietary
Speed (TPS)
Modalities
Input
text
textimageaudiovideo
Benchmarks sources: Cohere model card / Vellum LLM Leaderboard
GPQA Diamond 50 vendor 72 vendor
SWE-Bench Verified 25 vendor 55 vendor
Humanity's Last Exam 15 vendor 40.2
ARC-AGI 2 12 vendor 72.1
AIME 2025 45 vendor 78 vendor
MMMLU 74 vendor 85 vendor
BFCL 60 vendor 75 vendor
HumanEval 78 vendor 88 vendor
MATH 500 62 vendor 82 vendor
Avg benchmark score 70.1
Perf / dollar 20.8
Terminal-Bench 2.1 76.2
MCP Atlas 83.6
OSWorld-Verified 78.4
Independent composite scores each on its own scale — not part of the average above
LMSYS Chatbot Arena 1353.9 ELO 1476.4 ELO
AA Intelligence Index 52
AA Agentic Index 39.7
AA-Omniscience Index (−100–100) 21.2
Providers
Available from
Cohere — $2.50/$10.00/Mtok
Google — $1.50/$9.00/Mtok

Cost at scale

1M tokens · 50/50 input/output
Projected cost of Command A vs Gemini 3.5 Flash at increasing token volumes
VolumeCommand AGemini 3.5 FlashSavings
1M tokens $4.38 $3.38 $1 (22.9%)
10M tokens $43.75 $33.75 $10 (22.9%)
100M tokens $437.5 $337.5 $100 (22.9%)
1000M tokens $4375 $3375 $1000 (22.9%)

When to pick which

distilled from the pricing and spec data above
Pick Command A if…
  • this pairing gives Command A no edge on price, context, cache discount, or measured performance — pick it only for qualitative fit (ecosystem, compliance, model behavior on your prompts)
Pick Gemini 3.5 Flash if…
  • cost dominates: $3.38/Mtok blended vs $4.38 — 22.9% less on the same 50/50 token mix
  • you need the longer context: 1M tokens vs 256K (3.9×)
Try Gemini 3.5 Flash on Google

The cheaper option here — Gemini 3.5 Flash costs $3.38/Mtok blended on Google.

Get API key →

Summary

Command A by Cohere costs $2.50/Mtok input and $10.00/Mtok output, with a 256K-token context window. It supports text input.

Gemini 3.5 Flash by Google costs $1.50/Mtok input and $9.00/Mtok output, with a 1M-token context window. It supports text, image, audio, video input.

On a blended cost basis, Gemini 3.5 Flash is 22.9% cheaper than Command A.It also has a larger context window.

The two aren't directly comparable on average benchmark score: Gemini 3.5 Flash has published per-benchmark results, while Command A does not yet — it is measured today on independent composites (see the table above).

Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.

More comparisons

pairs sharing a model with this page