GPT-4
OpenAI's original GPT-4 model.
Key Specifications
- Context window
- 8K tokens
- Max output
- 8K tokens
- Knowledge cutoff
- Released
- Inputs
- Text
Hosted API pricing
| Provider | Input /1M tokens | Output /1M tokens | Cost at 10M in + 2M out | |
|---|---|---|---|---|
|
|
$30.00 | $60.00 | $420.00 | View |
|
|
$30.00 | $60.00 | $420.00 | View |
Heads up: Base-tier, on-demand rates per 1M tokens; cached, batch and long-context tiers excluded. A provider may serve a shorter context or a quantized build than the creator's release. Verify before provisioning. More on how we price.
Similarly priced models
The models nearest GPT-4 by blended rate, each at its own cheapest provider.
| Model | Blended /1M | Input /1M | Output /1M | Context | Cutoff | vs GPT-4 |
|---|---|---|---|---|---|---|
|
|
$16.6667 | $10.00 | $50.00 | 1M | Jan 2026 | −52% |
|
|
$16.6667 | $10.00 | $50.00 | 1M | Jun 2026 | −52% |
|
|
$16.6667 | $10.00 | $50.00 | 1M | Apr 2026 | −52% |
|
|
$22.50 | $15.00 | $60.00 | 200K | Oct 2023 | −36% |
|
|
$30.00 | $20.00 | $80.00 | 200K | Jun 2024 | −14% |
|
|
$32.50 | $15.00 | $120.00 | 400K | Sep 2024 | −7% |
|
|
$35.00 | $30.00 | $60.00 | 8K | Dec 2023 | |
|
|
$45.50 | $21.00 | $168.00 | 400K | Aug 2025 | +30% |
|
|
$55.00 | $30.00 | $180.00 | 1M | Aug 2025 | +57% |
|
|
$55.00 | $30.00 | $180.00 | 1M | Dec 2025 | +57% |
|
|
$225.00 | $150.00 | $600.00 | 200K | Oct 2023 | +543% |
Prices are USD per 1M tokens at each model's cheapest listed provider. Blended is the cost of 10M input plus 2M output tokens, spread over the 12M.
Frequently Asked Questions
What is GPT-4 good for?
Text-only work that needs the original GPT-4 for reproducibility. OpenAI calls it an older model and recommends newer ones for new work.
When is GPT-4 not a good fit?
Image input, tool calling and structured output, since it has none of them. It sits near the top of OpenAI's price range, and OpenAI plans to shut down the API model in October 2026.
What is the cheapest way to run GPT-4?
The cheapest hosted rate is $30.00 / $60.00 per 1M tokens via OpenAI.
How does GPT-4 compare with GPT-4 Turbo?
OpenAI calls GPT-4 Turbo the next generation of GPT-4, designed as a cheaper, more capable version of that model. GPT-4 lists at 200% more per 1M input tokens and 100% more per 1M output tokens than GPT-4 Turbo. It came out 13 months earlier, has a smaller context (8K against 128K tokens) and a larger max output (8K against 4K tokens) and takes no image input, which GPT-4 Turbo does.
Can I self-host GPT-4?
No. OpenAI does not publish the weights; GPT-4 is available only through hosted APIs.
More from OpenAI
| Model | Context | Input /1M | Output /1M |
|---|---|---|---|
|
|
400K | $15.00 | $120.00 |
|
|
200K | $20.00 | $80.00 |
|
|
400K | $21.00 | $168.00 |
|
|
200K | $15.00 | $60.00 |
|
|
1M | $30.00 | $180.00 |
|
|
1M | $30.00 | $180.00 |
|
|
1M | $10.00 | $50.00 |
|
|
200K | $10.00 | $40.00 |