Flagship models
All rates listed in USD per 1 Million (1M) tokens unless noted
StandardBatchFast mode
| Pricing Metric / Parameter | Standard RATE (per 1M) | Batch RATE (per 1M) | Fast mode RATE (per 1M) |
|---|---|---|---|
| Short context input | $0.40 | $0.20 | $0.70 |
| Short context cached input | $0.10 | - | $0.175 |
| Short context cache writes | - | - | - |
| Short context output | $1.60 | $0.80 | $2.80 |
| Long context input | - | - | - |
| Long context cached input | - | - | - |
| Long context cache writes | - | - | - |
| Long context output | - | - | - |
Source: OpenAI official pricing page, verified 2026-10-02. Prices in USD per 1M tokens.
LLM API Cost Comparison
Compare OpenAI model costs against Anthropic & Gemini
Use our free LLM Cost Calculator to benchmark gpt-4.1-mini token rates, Batch API discounts, and monthly estimated spend side-by-side with leading AI models.
Open LLM Cost Calculator →
Start with AWS