Alibaba Cloud logo

Qwen3.7-Max

The closed-weight flagship of the Qwen3.7 series, an agent-first reasoning model built for long autonomous sessions.

Cheapest of 2 providers, via Novita
Input
$1.25 / 1M tokens
Output
$3.75 / 1M tokens

About $20.00 for 10M input and 2M output tokens. Estimate yours

Novita Alibaba Cloud 2 providers

Key Specifications

Context window
1M tokens
Max output
66K tokens
Inputs
Text Outputs: Text

Qwen3.7-Max pricing by provider

Provider Input / 1M tokens Output / 1M tokens Cost at 10M in + 2M out
Novita logo Novita Cheapest $1.25 $3.75 $20.00 View
Alibaba Cloud logo Alibaba Cloud Creator $2.50 $7.50 $40.00 View

Heads up: We do our best to keep these specs & prices accurate. However, cloud costs may fluctuate based on region, usage, and other factors not listed here. These are estimates based on common setups and are for informational purposes only. Always verify current rates & exact specs with the provider before provisioning. LLM rates are base-tier, on-demand prices per 1M tokens; cached-input, batch and long-context tiers are not included.

Compare every model at this volume in the LLM cost calculator.

Capabilities

Function calling

Function calling

Connect to external tools, APIs, and systems.

Web search

Web search

Search the web for up-to-date information.

More from Alibaba Cloud

Model Context Input / 1M Output / 1M
Alibaba Cloud logo Qwen3-Coder-Plus 1M $1.00 $5.00
Alibaba Cloud logo Qwen3-VL-235B-A22B Thinking Open weights 262K $0.98 $3.95
Alibaba Cloud logo Qwen3-Max 262K $0.85 $3.38
Alibaba Cloud logo Qwen3.8-Max 1M $2.00 $6.00
Alibaba Cloud logo Qwen3.8-2.4T-A95B Open weights 262K $2.50 $6.25
Alibaba Cloud logo Qwen3.5-397B-A17B Open weights 262K $0.60 $3.60
Alibaba Cloud logo Qwen3.6-27B Open weights 262K $0.60 $3.60
Alibaba Cloud logo Qwen3.6-Plus 1M $0.50 $3.00

Models near this price

The models nearest this one by input rate, five each way, at each one's cheapest listed provider.

Frequently Asked Questions

How much does Qwen3.7-Max cost?

Qwen3.7-Max costs $1.25 per 1M input tokens and $3.75 per 1M output tokens via Novita, the cheapest of 2 providers listing it. The highest rate is Alibaba Cloud at $2.50 in / $7.50 out.

What does Qwen3.7-Max cost for 10M input and 2M output tokens?

At Novita's rates, 10M input tokens and 2M output tokens cost about $20.00: $12.50 for input and $7.50 for output. Input is prompts and context, output is what the model writes back; a workload that generates more than it reads shifts the cost toward the output rate of $3.75 per 1M tokens.

Which providers offer Qwen3.7-Max?

2 providers list Qwen3.7-Max: Novita ($1.25 in / $3.75 out) and Alibaba Cloud ($2.50 in / $7.50 out). Rates are per 1M tokens in USD, cheapest input rate first.

What is Qwen3.7-Max's context window?

Qwen3.7-Max accepts up to 1M tokens of input per request and returns up to 66K tokens per response. The context window is the prompt plus any documents, conversation history and tool results sent with it; every token in it is billed at the input rate.

What inputs and outputs does Qwen3.7-Max support?

Qwen3.7-Max accepts text as input and produces text. Its listed capabilities are function calling and web search.

How does Qwen3.7-Max compare with Qwen3-Coder-Plus?

Qwen3.7-Max costs $1.25 per 1M input tokens against Qwen3-Coder-Plus's $1.00, 1.2x more expensive ($3.75 vs $5.00 per 1M output tokens). Both have a 1M-token context window.

What are cheaper alternatives to Qwen3.7-Max?

Models from other creators with a lower input rate and at least Qwen3.7-Max's 1M-token context window: GLM-5.2 at $0.90 per 1M input tokens (1M context), Nemotron 3 Ultra 550B A55B at $0.60 per 1M input tokens (1M context) and Gemini 3 Flash Preview at $0.50 per 1M input tokens (1M context). Rates are the cheapest listed provider for each; whether the quality holds for a given task is a separate question.

Cheaper alternatives to Qwen3.7-Max