OpenAI

GPT-4o mini

The low-cost GPT-4o tier for high-volume, multimodal workloads at $0.15 / $0.60 per 1M.

ActiveReleased July 18, 2024Knowledge cutoff October 1, 2023Available from 4 providers
Tool callingCan invoke external tools and APIs (function calling) mid-response.Structured outputCan return JSON or schema-constrained responses instead of free text.ReasoningUses extended thinking / chain-of-thought for harder multi-step problems.Effort controlLets you dial thinking depth (low / medium / high) to trade cost for quality.Fine-tuningSupports training a custom variant on your own data (vendor fine-tune API or self-hosted open weights).

Source: OpenAI models, OpenAI pricing, Prompt caching (1,024 minimum, 30m TTL)

Price per 1M tokens
$0.15 in / $0.6 out
Context window128,000
Max output16,384
Cache read$0.075 / M
Batch discount-50%
Input
  • Text
  • Image
Output
  • Text
Priced across4 providers

Pricing

Every billing leg from the live catalog, in USD per 1M tokens.

Billing legRateBatch
Input$0.15 / M$0.075 / M
Output$0.6 / M$0.3 / M
Cache read$0.075 / Mn/a

Cost calculator

Billing-grade math over the rate card above, including the thinking tokens most estimates miss.

Estimated monthly cost
$42.00
at current list prices
Fresh input$30.00
Cached input$0.00
Output$12.00
Track this automatically →

Compare providers

The same model, priced across every platform that serves it. Lowest combined in/out rate is flagged.

PlatformRegionInput / MOutput / MCache read / MBatchAccess
OpenAILowest
-$0.15$0.6$0.075✓First-party API
Azure OpenAI
-$0.15$0.6$0.075✓Direct
Azure OpenAI · Data zone
-$0.165$0.66$0.083-Direct
Azure OpenAI · Regional
$0.165$0.66$0.083-Direct

Some hosts price by region. Shown rates match the selected region.

Price history

Has this model gotten cheaper or more expensive?

EffectiveInput / OutputStatus
April 20, 2026 – July 10, 2026$0.15 / $0.6past
July 10, 2026 – August 3, 2026$0.15 / $0.6past
August 3, 2026$0.15 / $0.6current

More from OpenAI

Other models from the same provider.

Stop estimating. Track it.

SuperPenguin meters every GPT-4o mini call your team makes, prices it with this exact rate card, and shows spend by feature, customer, and team.

Get started free

Rates are generated from SuperPenguin's live pricing catalog.