ModelPriceWatch.com
Last scan 2026-08-01 Models tracked 181 Providers 27 Cheapest paid Granite 4.0 Micro $0.017/Mtok in Every price links to its source

GPT-4.1 mini

by OpenAI

Current budget mid tier

Today's price · per 1M tokens

Input

$0.400

per 1M tokens

Output

$1.60

per 1M tokens

Blended

$1.00

avg of input & output $/1M

Cached input

$0.100

25% of input — prompt caching

Source: official OpenAI pricing · read Jul 4, 2026 MODELPRICEWATCH.COM · 2026-08-01

Overview

Compact, affordable model with 1M context. Good for high-throughput tasks.

Capabilities

struck through = not supported
Input 2/5
Text ✓ Image ✓ Audio Video PDF
Output 1/5
Text ✓ Image Audio Video Embedding
Features 3/9
Prompt caching ✓ Reasoning Coding Fast inference Long context ✓ Open weights Multimodal ✓ Web search Realtime

Benchmark performance

accuracy % · higher is better
Avg benchmark score55.7%
Perf per $/Mtok55.7
GPQA Diamond
65%
SWE-Bench Verified vendor
30%
Humanity's Last Exam vendor
12%
ARC-AGI 2 vendor
15%
AIME 2025 vendor
40%
MMMLU
78.5%
BFCL vendor
70%
HumanEval
23.6%
MATH 500 vendor
70%
Independent composite scores
  • LMSYS Chatbot Arena: 1383 ELO (Human preference)
How it stacks up
  • Ranks #55 of 68 benchmarked models by average score
  • Ranks #95 of 133 comparably-measured models by percentile score, across 4 independent measurements
  • Ranks #64 of 133 comparably-tested models by normalized performance per dollar
  • Strongest at MMMLU — 78.5%, #19 of 33

200 tokens/sec output

Source: Vellum LLM Leaderboard · updated Aug 1, 2026 · See full rankings →

Specifications

Provider
OpenAI
Context window
1M tokens
Modality
text, image
Parameters
Proprietary
Open source
No — proprietary
Released
Apr 14, 2025
Status
Current
Last updated
Jul 4, 2026
Tags
budgetmultimodal

Availability verified: Jul 25, 2026 — listed on OpenAI's own page