Flagship models
All rates listed in USD per 1 Million (1M) tokens unless noted
StandardBatchFlexFast mode
| Pricing Metric / Parameter | Standard RATE (per 1M) | Batch RATE (per 1M) | Flex RATE (per 1M) | Fast mode RATE (per 1M) |
|---|---|---|---|---|
| Short context input | $10.00 | $5.00 | $5.00 | $20.00 |
| Short context cached input | $1.00 | $0.50 | $0.50 | $2.00 |
| Short context cache writes | $12.50 | $6.25 | $6.25 | $25.00 |
| Short context output | $50.00 | $25.00 | $25.00 | $100.00 |
| Long context input | $20.00 | $10.00 | $10.00 | $40.00 |
| Long context cached input | $2.00 | $1.00 | $1.00 | $4.00 |
| Long context cache writes | $25.00 | $12.50 | $12.50 | $50.00 |
| Long context output | $75.00 | $37.50 | $37.50 | $150.00 |
Source: OpenAI official pricing page, verified 2026-09-21. Prices in USD per 1M tokens.
LLM API Cost Comparison
Compare OpenAI model costs against Anthropic & Gemini
Use our free LLM Cost Calculator to benchmark gpt-6-astra token rates, Batch API discounts, and monthly estimated spend side-by-side with leading AI models.
Open LLM Cost Calculator →
Start with AWS