OpenAI · LLM API Modelretiring

gpt-4o-2024-05-13

gpt-4o-2024-05-13 costs $5.00 per 1M input tokens and $15.00 per 1M output tokens. Batch API: $2.50 input, $7.50 output.

Updated September 21, 2026
Input / 1M
$5.00
Output / 1M
$15.00
Cached Input
—
Context Window
128K
Standard Input
$5.00 / 1M
Base prompt input rate
Standard Output
$15.00 / 1M
Completion output rate
Prompt Caching -50%
50% Off
Cached prompt context
Batch API -50%
$2.50 / 1M
Asynchronous 24h batch

Save up to 50% on OpenAI API token spend

Connect your OpenAI account — we'll optimize Batch API workflows and prompt caching efficiency.

OpenAI retires gpt-4o-2024-05-13 on 23 Oct 2026. Successor: gpt-5.6-sol.

Flagship models

All rates listed in USD per 1 Million (1M) tokens unless noted

StandardBatchFast mode
Pricing Metric / Parameter
Standard RATE
(per 1M)
Batch RATE
(per 1M)
Fast mode RATE
(per 1M)
Short context input$5.00$2.50$8.75
Short context cached input---
Short context cache writes---
Short context output$15.00$7.50$26.25
Long context input---
Long context cached input---
Long context cache writes---
Long context output---

Source: OpenAI official pricing page, verified 2026-09-21. Prices in USD per 1M tokens.

AI / LLM
LLM API Cost Comparison

Compare OpenAI model costs against Anthropic & Gemini

Use our free LLM Cost Calculator to benchmark gpt-4o-2024-05-13 token rates, Batch API discounts, and monthly estimated spend side-by-side with leading AI models.

Open LLM Cost Calculator →