GPT-5.4 Mini
Released March 2026 as the mid-size member of the GPT-5.4 family, aimed at coding, computer use and subagent workloads.
- Input
- $0.75 / 1M tokens
- Output
- $4.50 / 1M tokens
About $16.50 for 10M input and 2M output tokens. Estimate yours
Key Specifications
- Context window
- 400K tokens
- Max output
- 128K tokens
- Knowledge cutoff
- Inputs
- Text, Image Outputs: Text
GPT-5.4 Mini pricing by provider
| Provider | Input / 1M tokens | Output / 1M tokens | Cost at 10M in + 2M out | |
|---|---|---|---|---|
|
|
$0.75 | $4.50 | $16.50 | View |
|
|
$0.75 | $4.50 | $16.50 | View |
Heads up: We do our best to keep these specs & prices accurate. However, cloud costs may fluctuate based on region, usage, and other factors not listed here. These are estimates based on common setups and are for informational purposes only. Always verify current rates & exact specs with the provider before provisioning. LLM rates are base-tier, on-demand prices per 1M tokens; cached-input, batch and long-context tiers are not included.
Compare every model at this volume in the LLM cost calculator.
Capabilities
Function calling
Connect to external tools, APIs, and systems.
Structured output
Return responses in structured formats like JSON.
Web search
Search the web for up-to-date information.
Code execution
Write and execute code in a sandboxed environment.
File upload
Process uploaded documents, spreadsheets, and other files.
MCP
Connect to services via Model Context Protocol.
Image generation
Generate and edit images from text prompts.
More from OpenAI
| Model | Context | Input / 1M | Output / 1M |
|---|---|---|---|
|
|
200K | $1.00 | $4.00 |
|
|
200K | $1.10 | $4.40 |
|
|
16K | $0.50 | $1.50 |
|
|
400K | $1.25 | $10.00 |
|
|
400K | $1.25 | $10.00 |
|
|
1M | $0.40 | $1.60 |
|
|
400K | $1.75 | $14.00 |
|
|
400K | $1.75 | $14.00 |
Models near this price
The models nearest this one by input rate, five each way, at each one's cheapest listed provider.
Frequently Asked Questions
How much does GPT-5.4 Mini cost?
GPT-5.4 Mini costs $0.75 per 1M input tokens and $4.50 per 1M output tokens via OpenAI; all 2 providers listing it charge the same rate.
What does GPT-5.4 Mini cost for 10M input and 2M output tokens?
At OpenAI's rates, 10M input tokens and 2M output tokens cost about $16.50: $7.50 for input and $9.00 for output. Input is prompts and context, output is what the model writes back; a workload that generates more than it reads shifts the cost toward the output rate of $4.50 per 1M tokens.
Which providers offer GPT-5.4 Mini?
2 providers list GPT-5.4 Mini: OpenAI ($0.75 in / $4.50 out) and Azure ($0.75 in / $4.50 out). Rates are per 1M tokens in USD, cheapest input rate first.
What is GPT-5.4 Mini's context window?
GPT-5.4 Mini accepts up to 400K tokens of input per request and returns up to 128K tokens per response. The context window is the prompt plus any documents, conversation history and tool results sent with it; every token in it is billed at the input rate.
What is GPT-5.4 Mini's knowledge cutoff?
GPT-5.4 Mini's knowledge cutoff is August 2025: its training data runs up to that month and it has no built-in knowledge of later events. It supports web search, which can supply newer information at query time.
What inputs and outputs does GPT-5.4 Mini support?
GPT-5.4 Mini accepts text and images as input and produces text. Its listed capabilities are function calling, structured output, web search, code execution, file upload, MCP and image generation.
How does GPT-5.4 Mini compare with o4-mini?
GPT-5.4 Mini costs $0.75 per 1M input tokens against o4-mini's $1.00, 1.3x cheaper ($4.50 vs $4.00 per 1M output tokens). The context window is 400K tokens against 200K.
What are cheaper alternatives to GPT-5.4 Mini?
Models from other creators with a lower input rate and at least GPT-5.4 Mini's 400K-token context window: Nemotron 3 Ultra 550B A55B at $0.60 per 1M input tokens (1M context), Gemini 3 Flash Preview at $0.50 per 1M input tokens (1M context) and Qwen3.6-Plus at $0.50 per 1M input tokens (1M context). Rates are the cheapest listed provider for each; whether the quality holds for a given task is a separate question.
Cheaper alternatives to GPT-5.4 Mini
Nemotron 3 Ultra 550B A55B
$0.60 per 1M input tokens and $3.60 per 1M output tokens via Together, 1M-token context, by Nvidia.
Gemini 3 Flash Preview
$0.50 per 1M input tokens and $3.00 per 1M output tokens via Google Cloud, 1M-token context, by Google Cloud.
Qwen3.6-Plus
$0.50 per 1M input tokens and $3.00 per 1M output tokens via Alibaba Cloud, 1M-token context, by Alibaba Cloud.