ModelPriceWatch.com
Last scan 2026-08-09 Models tracked 198 Providers 30 Cheapest paid Granite 4.0 Micro $0.017/Mtok in Every price links to its source

Together

Hosting provider · 10 models tracked · Founded 2022

Inference platform hosting 200+ open-source models with simple per-token pricing. Strong coverage of Qwen, GLM, Kimi, MiniMax, and DeepSeek variants.

Together pricing at a glanceAugust 2026 · $ per 1M tokens

Together API pricing (August 2026): 10 current models range from $0.150 to $1.74 per 1M input tokens and $0.600 to $4.40 per 1M output tokens. The cheapest paid model is gpt-oss-120B at $0.150/1M input; the priciest is DeepSeek V4 Pro at $1.74/1M input / $3.48 output. Every price links to Together's official pricing page and refreshes twice daily.

Models10
Input range$0.150–$1.74
Output range$0.600–$4.40
Cached-input tiers6

Today's Together prices

10 models · cheapest blended first

Sorted by blended cost (cheapest first). Prices per 1M tokens, August 2026 — every price links to Together's official pricing page.

Current Together model prices per 1M tokens, sorted by blended cost
Model Blended* Input Output Cached in Relative cost Context Status
gpt-oss-120B
$0.262
$0.150 $0.600
128K Current
Gemma-4-31B-it-Pearl
$0.425
$0.280 $0.860
128K Current
MiniMax M3
$0.525
$0.300 $1.20 $0.060
1M Current
Qwen3.7-Plus
$0.560
$0.320 $1.28
128K Current
Llama 3.3 70B
$1.04
$1.04 $1.04
128K Current
NVIDIA Nemotron 3 Ultra
$1.35
$0.600 $3.60 $0.200
128K Current
Kimi K2.7 Code
$1.71
$0.950 $4.00 $0.190
256K Current
Qwen3.7-Max
$1.88
$1.25 $3.75 $0.130
1M Current
GLM-5.2
$2.15
$1.40 $4.40 $0.260
1M Current
DeepSeek V4 Pro
$2.17
$1.74 $3.48 $0.200
1M Current
* Blended = (3×input + 1×output) ÷ 4 $/1M tokens · cheap · mid · expensive MODELPRICEWATCH.COM · 2026-08-09

Quick stats

Models tracked
10
Type
hosting
Founded
2022
Cheapest model
gpt-oss-120B
Cheapest blended
$0.262/M

Open-weight models

8 open-weight models available from this provider.

GLM-5.2
$1.40/M in · 1M ctx
Gemma-4-31B-it-Pearl
$0.280/M in · 128K ctx
Kimi K2.7 Code
$0.950/M in · 256K ctx
Llama 3.3 70B
$1.04/M in · 128K ctx
MiniMax M3
$0.300/M in · 1M ctx
NVIDIA Nemotron 3 Ultra
$0.600/M in · 128K ctx

Try Together

Sign up and start building with Together models.

Get started
Run it yourself

Self-host Together's open-weight models

These models ship with open weights, so you can serve them yourself on rented GPUs instead of paying per-token API prices — often cheaper at scale.

Partner links — we may earn a commission if you sign up, at no cost to you. We only list platforms we'd recommend regardless.