ModelPriceWatch.com
Last scan 2026-08-09 Models tracked 198 Providers 30 Cheapest paid Granite 4.0 Micro $0.017/Mtok in Every price links to its source

LLM API providers

31 providers in the directory — commercial labs, open-source model makers, and inference hosting platforms. Model counts and cheapest prices come from the current price board.

All types 31
Commercial 25
Open source 1
Hosting 5

OpenAI

commercial

Maker of GPT and the OpenAI API. Offers frontier models (GPT-5.5, GPT-5.4), reasoning (o4-mini), and the open-source gpt-oss famil…

Models 24
Cheapest · avg $/1M $0.020

View all OpenAI models →

Anthropic

commercial

Maker of the Claude model family. Flagship Fable 5 and Opus 5, plus Sonnet and Haiku tiers. Known for safety research.

Models 11
Cheapest · avg $/1M $2.00

View all Anthropic models →

Google

commercial

Maker of Gemini multimodal models with up to 2M context. Current flagship Gemini 3.5 Flash; also open-weights Gemma.

Models 12
Cheapest · avg $/1M $0.175

View all Google models →

xAI

commercial

Maker of Grok models. Current flagship Grok 4.3 offers 1M context at $1.25/$2.50 per 1M tokens — strong value.

Models 4
Cheapest · avg $/1M $1.25

View all xAI models →

DeepSeek

commercial

Chinese AI lab offering extremely cheap APIs with aggressive prompt caching. V4 Flash at $0.14/$0.28 per 1M tokens.

Models 2
Cheapest · avg $/1M $0.175

View all DeepSeek models →

Mistral

commercial

European AI lab. Mistral Large 3 at $0.50/$1.50 per 1M tokens, plus Codestral for code and open-weight models.

Models 8
Cheapest · avg $/1M $0.100

View all Mistral models →

Amazon

commercial

Maker of the Nova model family on AWS. Nova Micro at $0.035/$0.14 per 1M tokens is among the cheapest hosted APIs.

Models 4
Cheapest · avg $/1M $0.061

View all Amazon models →

Cohere

commercial

Enterprise-focused lab. Command A flagship at $2.50/$10 per 1M tokens with a 256K context — the same list price as the older Comma…

Models 6
Cheapest · avg $/1M $0.020

View all Cohere models →

Meta

open

Maker of the open-weight Llama family. Llama 4 Scout (17B, 16 experts, 10M context) and Llama 3.3 70B are widely hosted.

Models 5
Cheapest · avg $/1M $0.058

View all Meta models →

Groq

hosting

Inference platform running open models on custom LPU chips for extreme speed. Up to 1000 tokens/sec, cheapest hosted prices for ma…

Models 6
Cheapest · avg $/1M $0.058

View all Groq models →

Together

hosting

Inference platform hosting 200+ open-source models with simple per-token pricing. Strong coverage of Qwen, GLM, Kimi, MiniMax, and…

Models 10
Cheapest · avg $/1M $0.262

View all Together models →

Fireworks

hosting

Inference platform serving open-weight models (Llama, Qwen, DeepSeek, Kimi, GLM, MiniMax, gpt-oss) on its own infrastructure at fi…

Models 15
Cheapest · avg $/1M $0.128

View all Fireworks models →

Perplexity

commercial

Search-augmented AI from Perplexity. The Sonar family (Sonar, Sonar Pro, Sonar Reasoning Pro, Sonar Deep Research) adds live web g…

Models 4
Cheapest · avg $/1M $1.00

View all Perplexity models →

NVIDIA

hosting

NVIDIA's Nemotron models for enterprise and agentic use — Llama Nemotron Ultra 253B and Nemotron 3 Ultra, from $0.60/$3.60 per 1M …

Models 0
Cheapest

View all NVIDIA models →

IBM

commercial

IBM's open-weight Granite 4 family for enterprise workloads — Granite 4 H Large down to Granite 4.0 Micro at $0.017/1M input token…

Models 5
Cheapest · avg $/1M $0.041

View all IBM models →

AI21 Labs

commercial

Israeli AI lab behind the Jamba hybrid SSM-Transformer models — Jamba Large and Jamba Mini ($0.20/$0.40 per 1M) with a 256K-token …

Models 2
Cheapest · avg $/1M $0.250

View all AI21 Labs models →

Reka

commercial

Multimodal AI lab. Reka Core, Flash, and Edge span frontier to on-device inference, with Reka Edge from $0.10/1M tokens.

Models 3
Cheapest · avg $/1M $0.100

View all Reka models →

Voyage AI

commercial

Voyage AI (a MongoDB company) builds high-accuracy embedding and reranking models for retrieval and RAG, billed per 1M input token…

Models 8
Cheapest · avg $/1M $0.020

View all Voyage AI models →

Alibaba

commercial

Alibaba's Qwen models — Qwen3-Max, Qwen-Plus, and Qwen-Turbo (from $0.05/$0.20 per 1M) with context windows up to 1M tokens.

Models 14
Cheapest · avg $/1M $0.055

View all Alibaba models →

Z.AI

commercial

Z.AI (formerly Zhipu AI) — makers of the GLM model family. Frontier and open-weight models with competitive pricing.

Models 14
Cheapest · avg $/1M $0

View all Z.AI models →

Moonshot

commercial

Chinese AI lab behind the Kimi models. Kimi K2.5 through K2.7 Code offer a 256K-token context from $0.60/1M input tokens.

Models 5
Cheapest · avg $/1M $1.20

View all Moonshot models →

MiniMax

commercial

Chinese AI lab. The MiniMax-M2 series (M2 through M2.7, plus M3 at 1M context) starts at $0.30/$1.20 per 1M tokens.

Models 2
Cheapest · avg $/1M $0.525

View all MiniMax models →

01.AI

commercial

Founded by Kai-Fu Lee. Yi Large is a bilingual English/Chinese model at $3/$9 per 1M tokens with a 32K context window.

Models 1
Cheapest · avg $/1M $4.50

View all 01.AI models →

Baichuan

commercial

Chinese AI lab. Baichuan M2-32B is an open-weight 32B model priced at $0.07/1M for both input and output tokens.

Models 1
Cheapest · avg $/1M $0.070

View all Baichuan models →

Tencent

commercial

Chinese technology conglomerate. Its Hunyuan family is served via Tencent Cloud's TokenHub API; Hunyuan Hy3 is a 295B-parameter Mo…

Models 1
Cheapest · avg $/1M $0.257

View all Tencent models →

DeepInfra

hosting

Serverless inference platform serving 100+ open-weight models (DeepSeek, Llama, Qwen, Gemma, Mistral) at pay-per-token rates — one…

Models 4
Cheapest · avg $/1M $0.113

View all DeepInfra models →

Meituan

commercial

Chinese technology company whose LongCat model family is served via the LongCat API Platform. LongCat-2.0 is offered pay-as-you-go…

Models 1
Cheapest · avg $/1M $1.30

View all Meituan models →

Sakana AI

commercial

Tokyo research company founded by David Ha, Llion Jones and Ren Ito, known for nature-inspired and evolutionary approaches to mode…

Models 1
Cheapest · avg $/1M $11.25

View all Sakana AI models →

Upstage

commercial

Korean AI company behind the Solar model family and a document-intelligence suite (Document Parse, Information Extract). Solar Pro…

Models 1
Cheapest · avg $/1M $0.262

View all Upstage models →

Relace

commercial

San Francisco lab building small, code-specific models for coding agents rather than general-purpose chat — retrieval, fast merge …

Models 2
Cheapest · avg $/1M $0.900

View all Relace models →

Cheapest = input & output averaged, $/1M tokens, for that provider's lowest-priced current model. Full board on the price index homepage. MODELPRICEWATCH.COM · 2026-08-09