ModelPriceWatch.com
Last scan 2026-09-22 Models tracked 267 Providers 35 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

LLM API providers

36 providers in the directory — commercial labs, open-source model makers, and inference hosting platforms. Model counts and cheapest prices come from the current price board.

All types 36
Commercial 29
Open source 1
Hosting 6

OpenAI

commercial

Maker of GPT and the OpenAI API. Offers frontier models (GPT-5.5, GPT-5.4), reasoning (o4-mini), and the open-source gpt-oss famil…

Models 40
Cheapest · avg $/1M $0.020

View all OpenAI models →

Anthropic

commercial

Maker of the Claude model family. Flagship Fable 5.1 and Opus 5.5, plus Sonnet and Haiku tiers. Known for safety research.

Models 14
Cheapest · avg $/1M $2.00

View all Anthropic models →

Google

commercial

Maker of Gemini multimodal models with up to 2M context. Current flagship Gemini 3.5 Flash; also open-weights Gemma.

Models 14
Cheapest · avg $/1M $0.175

View all Google models →

xAI

commercial

Maker of Grok models. Current flagship Grok 4.7 offers 500K context at $2.00/$6.00 per 1M tokens.

Models 6
Cheapest · avg $/1M $1.25

View all xAI models →

DeepSeek

commercial

Chinese AI lab offering extremely cheap APIs with aggressive prompt caching. From 2026-08-16 it bills peak/off-peak by the hour, w…

Models 2
Cheapest · avg $/1M $0.525

View all DeepSeek models →

Mistral

commercial

European AI lab. Mistral Large 3 at $0.50/$1.50 per 1M tokens, plus Codestral for code and open-weight models.

Models 7
Cheapest · avg $/1M $0.100

View all Mistral models →

Amazon

commercial

Maker of the Nova model family on AWS. Nova Micro at $0.035/$0.14 per 1M tokens is among the cheapest hosted APIs.

Models 3
Cheapest · avg $/1M $0.061

View all Amazon models →

Cohere

commercial

Enterprise-focused lab. Command A flagship at $2.50/$10 per 1M tokens with a 256K context — the same list price as the older Comma…

Models 5
Cheapest · avg $/1M $0.120

View all Cohere models →

Meta

open

Maker of the open-weight Llama family. Llama 4 Scout (17B, 16 experts, 10M context) and Llama 3.3 70B are widely hosted.

Models 6
Cheapest · avg $/1M $0.058

View all Meta models →

Groq

hosting

Inference platform running open models on custom LPU chips for extreme speed. Up to 1000 tokens/sec, cheapest hosted prices for ma…

Models 4
Cheapest · avg $/1M $0.131

View all Groq models →

Together

hosting

Inference platform hosting 200+ open-source models with simple per-token pricing. Strong coverage of Qwen, GLM, Kimi, MiniMax, and…

Models 10
Cheapest · avg $/1M $0.052

View all Together models →

Fireworks

hosting

Inference platform serving open-weight models (Llama, Qwen, DeepSeek, Kimi, GLM, MiniMax, gpt-oss) on its own infrastructure at fi…

Models 14
Cheapest · avg $/1M $0.128

View all Fireworks models →

Perplexity

commercial

Search-augmented AI from Perplexity. The Sonar family (Sonar, Sonar Pro, Sonar Reasoning Pro, Sonar Deep Research) adds live web g…

Models 4
Cheapest · avg $/1M $1.00

View all Perplexity models →

NVIDIA

hosting

NVIDIA's Nemotron model family for enterprise and agentic workloads. We do not currently track a first-party NVIDIA endpoint — Nem…

Models 0
Cheapest

View all NVIDIA models →

IBM

commercial

IBM's open-weight Granite 4 family for enterprise workloads — Granite 4 H Small down to Granite 4.0 Micro at $0.017/1M input token…

Models 3
Cheapest · avg $/1M $0.041

View all IBM models →

AI21 Labs

commercial

Israeli AI lab behind the Jamba hybrid SSM-Transformer models — Jamba Large and Jamba Mini ($0.20/$0.40 per 1M) with a 256K-token …

Models 2
Cheapest · avg $/1M $0.250

View all AI21 Labs models →

Reka

commercial

Multimodal AI lab. Reka Core, Flash, and Edge span frontier to on-device inference, with Reka Edge from $0.10/1M tokens.

Models 3
Cheapest · avg $/1M $0.100

View all Reka models →

Voyage AI

commercial

Voyage AI (a MongoDB company) builds high-accuracy embedding and reranking models for retrieval and RAG, billed per 1M input token…

Models 8
Cheapest · avg $/1M $0.020

View all Voyage AI models →

Alibaba

commercial

Alibaba's Qwen models — Qwen3-Max, Qwen-Plus, and Qwen-Turbo (from $0.05/$0.20 per 1M) with context windows up to 1M tokens.

Models 24
Cheapest · avg $/1M $0.055

View all Alibaba models →

Z.AI

commercial

Z.AI (formerly Zhipu AI) — makers of the GLM model family. Frontier and open-weight models with competitive pricing.

Models 23
Cheapest · avg $/1M $0

View all Z.AI models →

Moonshot

commercial

Chinese AI lab behind the Kimi models. Kimi K2.5 through K2.7 Code offer a 256K-token context from $0.60/1M input tokens.

Models 5
Cheapest · avg $/1M $1.20

View all Moonshot models →

MiniMax

commercial

Chinese AI lab. The MiniMax-M2 series (M2 through M2.7, plus M3 at 1M context) starts at $0.30/$1.20 per 1M tokens.

Models 2
Cheapest · avg $/1M $0.525

View all MiniMax models →

01.AI

commercial

Founded by Kai-Fu Lee. Its bilingual English/Chinese Yi models made 01.AI one of China's early frontier labs, but Yi Large, the mo…

Models 0
Cheapest

View all 01.AI models →

Baichuan

commercial

Chinese AI lab. Baichuan M2-32B is an open-weight 32B model priced at $0.07/1M for both input and output tokens.

Models 1
Cheapest · avg $/1M $0.070

View all Baichuan models →

Tencent

commercial

Chinese technology conglomerate. Its Hunyuan family is served via Tencent Cloud's TokenHub API; Hunyuan Hy3 is a 295B-parameter Mo…

Models 4
Cheapest · avg $/1M $0.077

View all Tencent models →

DeepInfra

hosting

Serverless inference platform serving 100+ open-weight models (DeepSeek, Llama, Qwen, Gemma, Mistral) at pay-per-token rates — one…

Models 6
Cheapest · avg $/1M $0.110

View all DeepInfra models →

Meituan

commercial

Chinese technology company whose LongCat model family is served via the LongCat API Platform. LongCat-2.0 is offered pay-as-you-go…

Models 1
Cheapest · avg $/1M $1.30

View all Meituan models →

Sakana AI

commercial

Tokyo research company founded by David Ha, Llion Jones and Ren Ito, known for nature-inspired and evolutionary approaches to mode…

Models 3
Cheapest · avg $/1M $1.71

View all Sakana AI models →

Upstage

commercial

Korean AI company behind the Solar model family and a document-intelligence suite (Document Parse, Information Extract). Solar Pro…

Models 2
Cheapest · avg $/1M $0.158

View all Upstage models →

Relace

commercial

San Francisco lab building small, code-specific models for coding agents rather than general-purpose chat — retrieval, fast merge …

Models 2
Cheapest · avg $/1M $0.900

View all Relace models →

Kwaipilot

commercial

Kuaishou's coding-model team, maker of the KwaiKAT / KAT-Coder family. Its API is sold first-party through StreamLake (溪流湖), Kuais…

Models 2
Cheapest · avg $/1M $0.259

View all Kwaipilot models →

Inception

commercial

Palo Alto lab building diffusion large language models (dLLMs), founded in 2024 by the Stanford, UCLA and Cornell researchers behi…

Models 3
Cheapest · avg $/1M $0.338

View all Inception models →

Unbiased

commercial

An AI platform built by Circuit & Chisel, the payments-and-identity-for-agents company founded by Louis Amira and David Noël-Romas…

Models 1
Cheapest · avg $/1M $3.75

View all Unbiased models →

TypeSafe AI

commercial

A San Francisco lab that deliberately took the opposite research direction to chat: instead of RLHF, it trains with what it calls …

Models 1
Cheapest · avg $/1M $0.032

View all TypeSafe AI models →

Cheapest = input & output averaged, $/1M tokens, for that provider's lowest-priced current model. Full board on the price index homepage. MODELPRICEWATCH.COM · 2026-09-22