OpenAI · LLM API Modelretiring

gpt-3.5-turbo-0125

gpt-3.5-turbo-0125 costs $0.50 per 1M input tokens and $1.50 per 1M output tokens. Batch API: $0.25 input, $0.75 output.

Updated September 21, 2026
Input / 1M
$0.50
Output / 1M
$1.50
Cached Input
—
Context Window
16K
Standard Input
$0.50 / 1M
Base prompt input rate
Standard Output
$1.50 / 1M
Completion output rate
Prompt Caching -50%
50% Off
Cached prompt context
Batch API -50%
$0.25 / 1M
Asynchronous 24h batch

Save up to 50% on OpenAI API token spend

Connect your OpenAI account — we'll optimize Batch API workflows and prompt caching efficiency.

OpenAI retires gpt-3.5-turbo-0125 on 23 Oct 2026. Successor: gpt-5.6-terra.

Flagship models

All rates listed in USD per 1 Million (1M) tokens unless noted

StandardBatch
Pricing Metric / Parameter
Standard RATE
(per 1M)
Batch RATE
(per 1M)
Short context input$0.50$0.25
Short context cached input--
Short context cache writes--
Short context output$1.50$0.75
Long context input--
Long context cached input--
Long context cache writes--
Long context output--

Source: OpenAI official pricing page, verified 2026-09-21. Prices in USD per 1M tokens.

AI / LLM
LLM API Cost Comparison

Compare OpenAI model costs against Anthropic & Gemini

Use our free LLM Cost Calculator to benchmark gpt-3.5-turbo-0125 token rates, Batch API discounts, and monthly estimated spend side-by-side with leading AI models.

Open LLM Cost Calculator →