OpenAI logo

GPT-3.5 Turbo

OpenAI's lightweight model of the GPT-3.5 generation.

List price via OpenAI
Input
$0.50 / 1M tokens
Output
$1.50 / 1M tokens
OpenAI Azure 2 providers

Key Specifications

Context window
16K tokens
Max output
4K tokens
Knowledge cutoff
Released
Inputs
Text

Hosted API pricing

Provider Input / 1M tokens Output / 1M tokens Cost at 10M in + 2M out
OpenAI logo OpenAI Creator $0.50 $1.50 $8.00 View
Microsoft Azure logo Azure $0.50 $1.50 $8.00 View

Heads up: Base-tier, on-demand rates per 1M tokens; cached, batch and long-context tiers excluded. A provider may serve a shorter context or a quantized build than the creator's release. Verify before provisioning. More on how we price.

Similarly priced models

The models nearest GPT-3.5 Turbo by blended rate, each at its own cheapest provider.

Model Blended / 1M vs GPT-3.5 Turbo
Alibaba Cloud Qwen3-235B-A22B Thinking (2507) Alibaba Cloud $0.575 −14%
Alibaba Cloud Qwen3-Coder-480B-A35B Alibaba Cloud $0.575 −14%
OpenAI GPT-4.1 Mini OpenAI $0.60 −10%
Google Cloud Gemini 2.5 Flash Google Cloud $0.6667 0%
Google Cloud Gemini 3.5 Flash-Lite Google Cloud $0.6667 0%
OpenAI GPT-3.5 Turbo This model OpenAI $0.6667
Mistral Mistral Large 3 Mistral $0.6667 0%
Alibaba Cloud Qwen3.5-Plus Alibaba Cloud $0.7333 +10%
Z.AI GLM-5 Z.AI $0.7667 +15%
DeepSeek DeepSeek R1 Distill Llama 70B DeepSeek $0.80 +20%
Z.AI GLM-4.7 Z.AI $0.8667 +30%

Prices are USD per 1M tokens at each model's cheapest listed provider. Blended is the cost of 10M input plus 2M output tokens, spread over the 12M.

Frequently Asked Questions

What is GPT-3.5 Turbo good for?

Cheap text-only chat and classification. OpenAI calls it a legacy model and has recommended GPT-4o Mini instead since July 2024.

When is GPT-3.5 Turbo not a good fit?

Images, tool calling and structured output, since it has none of them. Its knowledge is years out of date, and OpenAI plans to shut down the API model in October 2026.

What is the cheapest way to run GPT-3.5 Turbo?

The cheapest hosted rate is $0.50 / $1.50 per 1M tokens via OpenAI.

How does GPT-3.5 Turbo compare with GPT-4o Mini?

OpenAI recommends GPT-4o Mini in place of GPT-3.5 Turbo and says it is smarter, cheaper and just as fast, with better long-context performance, and can be fine-tuned or distilled from GPT-4o. GPT-3.5 Turbo lists at 233% more per 1M input tokens and 150% more per 1M output tokens than GPT-4o Mini. It came out 16 months earlier, has an older knowledge cutoff (September 2021 against October 2023), a smaller context (16K against 128K tokens) and a smaller max output (4K against 16K tokens) and takes no image input, which GPT-4o Mini does.

Can I self-host GPT-3.5 Turbo?

No. OpenAI does not publish the weights; GPT-3.5 Turbo is available only through hosted APIs.

More from OpenAI

Model Context Input / 1M Output / 1M
OpenAI GPT-4.1 Mini 1M $0.40 $1.60
OpenAI GPT-5 Mini 400K $0.25 $2.00
OpenAI GPT-5.4 Nano 400K $0.20 $1.25
OpenAI GPT-5.6 Luna 1M $0.20 $1.20
OpenAI GPT-5.4 Mini 400K $0.75 $4.50
OpenAI o4-mini 200K $1.00 $4.00
OpenAI o3-mini 200K $1.10 $4.40
OpenAI GPT-4o Mini 128K $0.15 $0.60