Qwen3-Coder-Plus
The hosted proprietary tier of the Qwen3-Coder family, released alongside the open Qwen3-Coder-480B-A35B and built for agentic coding through tool calling and environment interaction.
- Input
- $1.00 / 1M tokens
- Output
- $5.00 / 1M tokens
About $20.00 for 10M input and 2M output tokens. Estimate yours
Key Specifications
- Context window
- 1M tokens
- Max output
- 66K tokens
- Inputs
- Text Outputs: Text
Qwen3-Coder-Plus pricing by provider
| Provider | Input / 1M tokens | Output / 1M tokens | Cost at 10M in + 2M out | |
|---|---|---|---|---|
|
|
$1.00 | $5.00 | $20.00 | View |
Heads up: We do our best to keep these specs & prices accurate. However, cloud costs may fluctuate based on region, usage, and other factors not listed here. These are estimates based on common setups and are for informational purposes only. Always verify current rates & exact specs with the provider before provisioning. LLM rates are base-tier, on-demand prices per 1M tokens; cached-input, batch and long-context tiers are not included.
Compare every model at this volume in the LLM cost calculator.
Capabilities
Function calling
Connect to external tools, APIs, and systems.
More from Alibaba Cloud
| Model | Context | Input / 1M | Output / 1M |
|---|---|---|---|
| Qwen3-VL-235B-A22B Thinking | 262K | $0.98 | $3.95 |
| Qwen3-Max | 262K | $0.85 | $3.38 |
| Qwen3.7-Max | 1M | $1.25 | $3.75 |
| Qwen3.5-397B-A17B | 262K | $0.60 | $3.60 |
| Qwen3.6-27B | 262K | $0.60 | $3.60 |
| Qwen3.6-Plus | 1M | $0.50 | $3.00 |
| Qwen3.8-Max | 1M | $2.00 | $6.00 |
| Qwen3.5-Plus | 1M | $0.40 | $2.40 |
Cheaper alternatives to Qwen3-Coder-Plus
GLM-5.2
$0.90 per 1M input tokens and $2.80 per 1M output tokens via Geodd, 1M-token context, by Z.AI.
Nemotron 3 Ultra 550B A55B
$0.60 per 1M input tokens and $3.60 per 1M output tokens via Together, 1M-token context, by Nvidia.
Gemini 3 Flash Preview
$0.50 per 1M input tokens and $3.00 per 1M output tokens via Google Cloud, 1M-token context, by Google Cloud.
Models near this price
The models nearest this one by input rate, five each way, at each one's cheapest listed provider.
Frequently Asked Questions
How much does Qwen3-Coder-Plus cost?
Qwen3-Coder-Plus costs $1.00 per 1M input tokens and $5.00 per 1M output tokens via Alibaba Cloud.
What does Qwen3-Coder-Plus cost for 10M input and 2M output tokens?
At Alibaba Cloud's rates, 10M input tokens and 2M output tokens cost about $20.00: $10.00 for input and $10.00 for output. Input is prompts and context, output is what the model writes back; a workload that generates more than it reads shifts the cost toward the output rate of $5.00 per 1M tokens.
Which providers offer Qwen3-Coder-Plus?
One provider lists Qwen3-Coder-Plus: Alibaba Cloud ($1.00 in / $5.00 out).
What is Qwen3-Coder-Plus's context window?
Qwen3-Coder-Plus accepts up to 1M tokens of input per request and returns up to 66K tokens per response. The context window is the prompt plus any documents, conversation history and tool results sent with it; every token in it is billed at the input rate.
What inputs and outputs does Qwen3-Coder-Plus support?
Qwen3-Coder-Plus accepts text as input and produces text. Its listed capabilities are function calling.
How does Qwen3-Coder-Plus compare with Qwen3-VL-235B-A22B Thinking?
Qwen3-Coder-Plus costs $1.00 per 1M input tokens against Qwen3-VL-235B-A22B Thinking's $0.98, 1x more expensive ($5.00 vs $3.95 per 1M output tokens). The context window is 1M tokens against 262K.
What are cheaper alternatives to Qwen3-Coder-Plus?
Models from other creators with a lower input rate and at least Qwen3-Coder-Plus's 1M-token context window: GLM-5.2 at $0.90 per 1M input tokens (1M context), Nemotron 3 Ultra 550B A55B at $0.60 per 1M input tokens (1M context) and Gemini 3 Flash Preview at $0.50 per 1M input tokens (1M context). Rates are the cheapest listed provider for each; whether the quality holds for a given task is a separate question.