Today's price · per 1M tokens
Input
$0.100
per 1M tokens
Output
$0.300
per 1M tokens
Blended
$0.150
blended $/1M — 3:1 weighted input:output
Source: official Mistral pricing · read Jun 25, 2026
MODELPRICEWATCH.COM · 2026-08-09
Overview
Speech-to-text and audio understanding model. $0.10/$0.30 per 1M tokens.
Run it yourself
Deploy this open model on rented GPUs
Open weights mean you can self-host instead of paying per-token API prices. These platforms let you serve it on demand — often cheaper at scale.
RunPod →
Rent GPUs by the second and serve any open-weight model yourself.
Novita AI →
Serverless open-model inference billed per token — no GPU ops.
Vast.ai →
GPU marketplace — rent spot/on-demand GPUs cheaply and serve any open-weight model.
Partner links — we may earn a commission if you sign up, at no cost to you. We only list platforms we'd recommend regardless.
Capabilities
struck through = not supportedInput 2/5
Text ✓
Image
Audio ✓
Video
PDF
Output 1/5
Text ✓
Image
Audio
Video
Embedding
Features 3/9
Prompt caching
Reasoning
Coding
Fast inference
Long context ✓
Open weights ✓
Multimodal ✓
Web search
Realtime
Specifications
- Provider
- Mistral
- Context window
- 128K tokens
- Modality
- text, audio
- Parameters
- 24B
- Open source
- Yes — open weights available
- Released
- Jul 1, 2025
- Status
- Current
- Last updated
- Jun 25, 2026
- Tags
Availability verified: Aug 2, 2026 — listed on Mistral's own page