OpenAI · price per token

GPT-4o mini pricing

What GPT-4o mini costs per million tokens: input, output, cache read and cache write, with batch and long-context rates where the provider publishes them. Every figure is the rate Vigil's proxy bills with, read from the provider's own page.

Prices checked 9 Oct 2026 against developers.openai.com. All shown rates agreed with the page.

GPT-4o mini on OpenAI API

Model id gpt-4o-mini. Verified against the provider's published price; registry row retrieved 2026-09-30 from developers.openai.com.

AppliesInputOutputCache readCache write
Standard$0.15$0.60$0.075—

USD per 1M tokens. A dash means the provider publishes no rate for that dimension, so the proxy declines to price it.

  • Endpoint classes: global, regional.
  • on_demand: the standard rate.
  • batch: 50% below the standard rate on input and output; cache rates unpublished.
  • No long-context tier: the standard rate applies at any prompt size.

Price a GPT-4o mini call

GPT-4o mini on OpenAI API, standard rate, priced now

Per call$0.001980

Cache writes, batch, endpoint class and monthly volume: AI token calculator

Questions

What does a cache read cost on GPT-4o mini?
A cache-read token on OpenAI API costs $0.075 per 1M tokens, against $0.15 for an ordinary input token. An ordinary input token costs 2 times as much. No cache-write rate is published for this model.
Does GPT-4o mini have a batch price?
Yes: batch: 50% below the standard rate on input and output; cache rates unpublished. The batch rate applies to the whole request; the proxy prices a batch call from the same row with that factor applied.

Other models: Claude Opus 5 · Claude Sonnet 5 · Claude Haiku 4.5 · Claude Fable 5.1 · Claude Opus 4.8 · Claude Sonnet 4.6 · GPT-6 Sol · GPT-6 Luna · GPT-5.6 Sol · GPT-5.6 Terra · GPT-5.3 Codex · GPT-4o · Gemini 3.1 Pro · Gemini 3.5 Flash · Gemini 3.8 Flash · DeepSeek V4 Pro · DeepSeek Flash · Grok 4.6 · Kimi K3