Mistral logo

Mistral Small 4

A 119B open-weight MoE under Apache 2.0 with 128 experts and roughly 6B active parameters per token, unifying Mistral Small, Magistral and Devstral into one hybrid model.

Price via Mistral
Input
$0.15 / 1M tokens
Output
$0.60 / 1M tokens

About $2.70 for 10M input and 2M output tokens. Estimate yours

Mistral 1 provider

Key Specifications

Context window
256K tokens
Inputs
Text, Image Outputs: Text

Mistral Small 4 pricing by provider

Provider Input / 1M tokens Output / 1M tokens Cost at 10M in + 2M out
Mistral logo Mistral Creator $0.15 $0.60 $2.70 View

Heads up: We do our best to keep these specs & prices accurate. However, cloud costs may fluctuate based on region, usage, and other factors not listed here. These are estimates based on common setups and are for informational purposes only. Always verify current rates & exact specs with the provider before provisioning. LLM rates are base-tier, on-demand prices per 1M tokens; cached-input, batch and long-context tiers are not included.

Compare every model at this volume in the LLM cost calculator.

Capabilities

Function calling

Function calling

Connect to external tools, APIs, and systems.

Structured output

Structured output

Return responses in structured formats like JSON.

More from Mistral

Model Context Input / 1M Output / 1M
Mistral logo Ministral 3 8B 256K $0.15 $0.15
Mistral logo Ministral 3 14B 256K $0.20 $0.20
Mistral logo Ministral 3 3B 256K $0.10 $0.10
Mistral logo Codestral 128K $0.30 $0.90
Mistral logo Mistral Large 3 256K $0.50 $1.50
Mistral logo Mistral NeMo 128K $0.04 $0.17
Mistral logo Mistral Medium 3.5 256K $1.50 $7.50
Mistral logo Devstral Medium 128K

Models near this price

The models nearest this one by input rate, five each way, at each one's cheapest listed provider.

Frequently Asked Questions

How much does Mistral Small 4 cost?

Mistral Small 4 costs $0.15 per 1M input tokens and $0.60 per 1M output tokens via Mistral.

What does Mistral Small 4 cost for 10M input and 2M output tokens?

At Mistral's rates, 10M input tokens and 2M output tokens cost about $2.70: $1.50 for input and $1.20 for output. Input is prompts and context, output is what the model writes back; a workload that generates more than it reads shifts the cost toward the output rate of $0.60 per 1M tokens.

Which providers offer Mistral Small 4?

One provider lists Mistral Small 4: Mistral ($0.15 in / $0.60 out).

What is Mistral Small 4's context window?

Mistral Small 4 accepts up to 256K tokens of input per request. The context window is the prompt plus any documents, conversation history and tool results sent with it; every token in it is billed at the input rate.

What inputs and outputs does Mistral Small 4 support?

Mistral Small 4 accepts text and images as input and produces text. Its listed capabilities are function calling and structured output.

How does Mistral Small 4 compare with Ministral 3 8B?

Mistral Small 4 and Ministral 3 8B both cost $0.15 per 1M input tokens ($0.60 vs $0.15 per 1M output tokens). Both have a 256K-token context window.

What are cheaper alternatives to Mistral Small 4?

Models from other creators with a lower input rate and at least Mistral Small 4's 256K-token context window: DeepSeek V4 Flash at $0.14 per 1M input tokens (1M context), Hy3 at $0.14 per 1M input tokens (262K context) and Gemma 4 26B A4B at $0.13 per 1M input tokens (256K context). Rates are the cheapest listed provider for each; whether the quality holds for a given task is a separate question.

Cheaper alternatives to Mistral Small 4