Today's price · per 1M tokens
Input
$0.100
per 1M tokens
Output
$0.300
per 1M tokens
Blended
$0.150
blended $/1M — 3:1 weighted input:output
Cached input
$0.020
20% of input — prompt caching
Price receipt
We read StepFun's own pricing page on and found step-3.5-flash listed with a price on the same row — that read is where the number above comes from, recorded as a new model. We keep the page text we read; its content hash is ca8a14ccf4.
Overview
StepFun's text-only Flash reasoning model, a 196B-parameter sparse MoE with 11B active per token and a 256K-token context, built for fast tool calling and multi-step agent tasks. StepFun released the weights under Apache 2.0 and sells it on its own Open Platform, alongside an agent-tuned snapshot, step-3.5-flash-2603, at the same rate. $0.10/$0.30 per 1M tokens.
Deploy this open model on rented GPUs
Open weights mean you can self-host instead of paying per-token API prices. These platforms let you serve it on demand — often cheaper at scale.
Partner links — we may earn a commission if you sign up, at no cost to you. We only list platforms we'd recommend regardless.
Capabilities
struck through = not supportedSpecifications
- Provider
- StepFun
- Context window
- 262K tokens
- Modality
- text
- Parameters
- 196B (A11B)
- Open source
- Yes — open weights available
- Released
- Status
- Current
- Last updated
- Tags
Availability verified: — listed on StepFun's own page