Today's price · per 1M tokens
Input
$0.400
per 1M tokens
Output
$1.60
per 1M tokens
Blended
$1.00
avg of input & output $/1M
Cached input
$0.100
25% of input — prompt caching
Source: official OpenAI pricing · read Jul 4, 2026
MODELPRICEWATCH.COM · 2026-08-01
Overview
Compact, affordable model with 1M context. Good for high-throughput tasks.
Capabilities
struck through = not supportedInput 2/5
Text ✓
Image ✓
Audio
Video
PDF
Output 1/5
Text ✓
Image
Audio
Video
Embedding
Features 3/9
Prompt caching ✓
Reasoning
Coding
Fast inference
Long context ✓
Open weights
Multimodal ✓
Web search
Realtime
Benchmark performance
accuracy % · higher is betterAvg benchmark score55.7%
Perf per $/Mtok55.7
GPQA Diamond
65%
SWE-Bench Verified vendor
30%
Humanity's Last Exam vendor
12%
ARC-AGI 2 vendor
15%
AIME 2025 vendor
40%
MMMLU
78.5%
BFCL vendor
70%
HumanEval
23.6%
MATH 500 vendor
70%
Independent composite scores
- LMSYS Chatbot Arena: 1383 ELO (Human preference)
How it stacks up
- Ranks #55 of 68 benchmarked models by average score
- Ranks #95 of 133 comparably-measured models by percentile score, across 4 independent measurements
- Ranks #64 of 133 comparably-tested models by normalized performance per dollar
- Strongest at MMMLU — 78.5%, #19 of 33
200 tokens/sec output
Source: Vellum LLM Leaderboard · updated Aug 1, 2026 · See full rankings →
Specifications
- Provider
- OpenAI
- Context window
- 1M tokens
- Modality
- text, image
- Parameters
- Proprietary
- Open source
- No — proprietary
- Released
- Apr 14, 2025
- Status
- Current
- Last updated
- Jul 4, 2026
- Tags
Availability verified: Jul 25, 2026 — listed on OpenAI's own page