OpenAI · LLM API Modelretiring

gpt-3.5-turbo-instruct

gpt-3.5-turbo-instruct costs $1.50 per 1M input tokens and $2.00 per 1M output tokens.

Updated September 21, 2026
Input / 1M
$1.50
Output / 1M
$2.00
Cached Input
—
Context Window
8K
Standard Input
$1.50 / 1M
Base prompt input rate
Standard Output
$2.00 / 1M
Completion output rate
Prompt Caching -50%
50% Off
Cached prompt context
Batch API -50%
50% Off
Asynchronous 24h batch

Save up to 50% on OpenAI API token spend

Connect your OpenAI account — we'll optimize Batch API workflows and prompt caching efficiency.

OpenAI retires gpt-3.5-turbo-instruct on 28 Sep 2026. Successor: gpt-5.6-terra.

Flagship models

All rates listed in USD per 1 Million (1M) tokens unless noted

Standard
Pricing Metric / Parameter
Standard RATE
(per 1M)
Short context input$1.50
Short context cached input-
Short context cache writes-
Short context output$2.00
Long context input-
Long context cached input-
Long context cache writes-
Long context output-

Source: OpenAI official pricing page, verified 2026-09-21. Prices in USD per 1M tokens.

AI / LLM
LLM API Cost Comparison

Compare OpenAI model costs against Anthropic & Gemini

Use our free LLM Cost Calculator to benchmark gpt-3.5-turbo-instruct token rates, Batch API discounts, and monthly estimated spend side-by-side with leading AI models.

Open LLM Cost Calculator →