DeepSeek V4 Flash Vision Exp
Released August 2026 as an experimental multimodal variant of DeepSeek V4 Flash that adds image input over the same text capabilities in agents, reasoning and world knowledge.
- Input
- $0.44 / 1M tokens
- Output
- $1.32 / 1M tokens
About $7.04 for 10M input and 2M output tokens. Estimate yours
Key Specifications
- Context window
- 1M tokens
- Max output
- 384K tokens
- Inputs
- Text, Image Outputs: Text
DeepSeek V4 Flash Vision Exp pricing by provider
| Provider | Input / 1M tokens | Output / 1M tokens | Cost at 10M in + 2M out | |
|---|---|---|---|---|
|
|
$0.44 | $1.32 | $7.04 | View |
Heads up: We do our best to keep these specs & prices accurate. However, cloud costs may fluctuate based on region, usage, and other factors not listed here. These are estimates based on common setups and are for informational purposes only. Always verify current rates & exact specs with the provider before provisioning. LLM rates are base-tier, on-demand prices per 1M tokens; cached-input, batch and long-context tiers are not included.
Compare every model at this volume in the LLM cost calculator.
Capabilities
Function calling
Connect to external tools, APIs, and systems.
Structured output
Return responses in structured formats like JSON.
More from DeepSeek
| Model | Context | Input / 1M | Output / 1M |
|---|---|---|---|
|
|
1M | $0.45 | $0.89 |
|
|
64K | $0.40 | $1.30 |
|
|
64K | $0.70 | $2.50 |
|
|
131K | $0.27 | $1.00 |
|
|
131K | $0.27 | $1.00 |
|
|
131K | $0.27 | $0.40 |
|
|
128K | $0.27 | $1.12 |
|
|
131K | $0.80 | $0.80 |
Models near this price
The models nearest this one by input rate, five each way, at each one's cheapest listed provider.
Frequently Asked Questions
How much does DeepSeek V4 Flash Vision Exp cost?
DeepSeek V4 Flash Vision Exp costs $0.44 per 1M input tokens and $1.32 per 1M output tokens via DeepSeek.
What does DeepSeek V4 Flash Vision Exp cost for 10M input and 2M output tokens?
At DeepSeek's rates, 10M input tokens and 2M output tokens cost about $7.04: $4.40 for input and $2.64 for output. Input is prompts and context, output is what the model writes back; a workload that generates more than it reads shifts the cost toward the output rate of $1.32 per 1M tokens.
Which providers offer DeepSeek V4 Flash Vision Exp?
One provider lists DeepSeek V4 Flash Vision Exp: DeepSeek ($0.44 in / $1.32 out).
What is DeepSeek V4 Flash Vision Exp's context window?
DeepSeek V4 Flash Vision Exp accepts up to 1M tokens of input per request and returns up to 384K tokens per response. The context window is the prompt plus any documents, conversation history and tool results sent with it; every token in it is billed at the input rate.
What inputs and outputs does DeepSeek V4 Flash Vision Exp support?
DeepSeek V4 Flash Vision Exp accepts text and images as input and produces text. Its listed capabilities are function calling and structured output.
How does DeepSeek V4 Flash Vision Exp compare with DeepSeek V4 Pro?
DeepSeek V4 Flash Vision Exp costs $0.44 per 1M input tokens against DeepSeek V4 Pro's $0.45, 1x cheaper ($1.32 vs $0.89 per 1M output tokens). Both have a 1M-token context window.
What are cheaper alternatives to DeepSeek V4 Flash Vision Exp?
Models from other creators with a lower input rate and at least DeepSeek V4 Flash Vision Exp's 1M-token context window: GPT-4.1 Mini at $0.40 per 1M input tokens (1M context), Qwen3.5-Plus at $0.40 per 1M input tokens (1M context) and Qwen3.7-Plus at $0.40 per 1M input tokens (1M context). Rates are the cheapest listed provider for each; whether the quality holds for a given task is a separate question.
Cheaper alternatives to DeepSeek V4 Flash Vision Exp
GPT-4.1 Mini
$0.40 per 1M input tokens and $1.60 per 1M output tokens via OpenAI, 1M-token context, by OpenAI.
Qwen3.5-Plus
$0.40 per 1M input tokens and $2.40 per 1M output tokens via Alibaba Cloud, 1M-token context, by Alibaba Cloud.
Qwen3.7-Plus
$0.40 per 1M input tokens and $1.60 per 1M output tokens via Alibaba Cloud, 1M-token context, by Alibaba Cloud.