ModelPriceWatch.com
Last scan 2026-09-22 Models tracked 267 Providers 35 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

Together

Hosting provider · 10 models tracked · Founded 2022

Inference platform hosting 200+ open-source models with simple per-token pricing. Strong coverage of Qwen, GLM, Kimi, MiniMax, and DeepSeek variants.

Together pricing at a glanceSeptember 2026 · $ per 1M tokens

Together API pricing (September 2026): 10 current models range from $0.030 to $2.50 per 1M input tokens and $0.120 to $7.50 per 1M output tokens. The cheapest paid model is LFM2.5 8B A1B at $0.030/1M input; the priciest is Qwen3.7-Max at $2.50/1M input / $7.50 output. Every price links to Together's official pricing page and refreshes twice daily.

Models10
Input range$0.030–$2.50
Output range$0.120–$7.50
Cached-input tiers6

Today's Together prices

10 models · cheapest blended first

Sorted by blended cost (cheapest first). Prices per 1M tokens, September 2026 — every price links to Together's official pricing page.

Current Together model prices per 1M tokens, sorted by blended cost
Model Blended* Input Output Cached in Relative cost Context Status
LFM2.5 8B A1B
$0.052
$0.030 $0.120
33K Current
GPT-OSS 120B
$0.262
$0.150 $0.600
131K Current
MiniMax-M3
$0.525
$0.300 $1.20 $0.060
524K Current
Qwen3.7-Plus
$0.560
$0.320 $1.28
1M Current
Muse Glimmer 30B
$0.637
$0.350 $1.50 $0.040
131K Current
Llama 3.3 70B
$1.04
$1.04 $1.04
131K Current
DeepSeek V4 Pro
$1.98
$1.32 $3.96 $0.130
1M Current
GLM-5.2
$2.15
$1.40 $4.40 $0.260
1M Current
GLM-5.3
$2.15
$1.40 $4.40 $0.260
1M Current
Qwen3.7-Max
$3.75
$2.50 $7.50 $0.250
1M Current
* Blended = (3×input + 1×output) ÷ 4 $/1M tokens · cheap · mid · expensive MODELPRICEWATCH.COM · 2026-09-22

Quick stats

Models tracked
10
Type
hosting
Founded
2022
Cheapest model
LFM2.5 8B A1B
Cheapest blended
$0.052/M

Open-weight models

8 open-weight models available from this provider.

GLM-5.2
$1.40/M in · 1M ctx
GLM-5.3
$1.40/M in · 1M ctx
Muse Glimmer 30B
$0.350/M in · 131K ctx
LFM2.5 8B A1B
$0.030/M in · 33K ctx
Llama 3.3 70B
$1.04/M in · 131K ctx
MiniMax-M3
$0.300/M in · 524K ctx

Try Together

Sign up and start building with Together models.

Get started
Run it yourself

Self-host Together's open-weight models

These models ship with open weights, so you can serve them yourself on rented GPUs instead of paying per-token API prices — often cheaper at scale.

Partner links — we may earn a commission if you sign up, at no cost to you. We only list platforms we'd recommend regardless.