Z.AI logo

GLM-5V-Turbo

Released April 2026, Z.AI's first natively multimodal coding model, pairing a CogViT vision encoder with the GLM-5-Turbo base for design-to-code, screenshot debugging and GUI agent tasks.

Cheapest of 2 providers, via Z.AI
Input
$1.20 / 1M tokens
Output
$4.00 / 1M tokens

About $20.00 for 10M input and 2M output tokens. Estimate yours

Z.AI Novita 2 providers

Key Specifications

Context window
205K tokens
Max output
131K tokens
Inputs
Text, Image, Video Outputs: Text

GLM-5V-Turbo pricing by provider

Provider Input / 1M tokens Output / 1M tokens Cost at 10M in + 2M out
Z.AI logo Z.AI Creator Cheapest $1.20 $4.00 $20.00 View
Novita logo Novita $1.20 $4.00 $20.00 View

Heads up: We do our best to keep these specs & prices accurate. However, cloud costs may fluctuate based on region, usage, and other factors not listed here. These are estimates based on common setups and are for informational purposes only. Always verify current rates & exact specs with the provider before provisioning. LLM rates are base-tier, on-demand prices per 1M tokens; cached-input, batch and long-context tiers are not included.

Compare every model at this volume in the LLM cost calculator.

Capabilities

Function calling

Function calling

Connect to external tools, APIs, and systems.

More from Z.AI

Model Context Input / 1M Output / 1M
Z.AI logo GLM-5-Turbo 205K $1.20 $4.00
Z.AI logo GLM-5.1 Open weights 205K $1.38 $4.40
Z.AI logo GLM-5.3 1M $1.40 $4.40
Z.AI logo GLM-5.2 Open weights 1M $0.90 $2.80
Z.AI logo GLM-4.7 Open weights 205K $0.60 $2.20
Z.AI logo GLM-5 Open weights 200K $0.60 $1.60
Z.AI logo GLM-5.3-Flash Open weights 1M $0.15 $0.50

Models near this price

The models nearest this one by input rate, five each way, at each one's cheapest listed provider.

Frequently Asked Questions

How much does GLM-5V-Turbo cost?

GLM-5V-Turbo costs $1.20 per 1M input tokens and $4.00 per 1M output tokens via Z.AI; all 2 providers listing it charge the same rate.

What does GLM-5V-Turbo cost for 10M input and 2M output tokens?

At Z.AI's rates, 10M input tokens and 2M output tokens cost about $20.00: $12.00 for input and $8.00 for output. Input is prompts and context, output is what the model writes back; a workload that generates more than it reads shifts the cost toward the output rate of $4.00 per 1M tokens.

Which providers offer GLM-5V-Turbo?

2 providers list GLM-5V-Turbo: Z.AI ($1.20 in / $4.00 out) and Novita ($1.20 in / $4.00 out). Rates are per 1M tokens in USD, cheapest input rate first.

What is GLM-5V-Turbo's context window?

GLM-5V-Turbo accepts up to 205K tokens of input per request and returns up to 131K tokens per response. The context window is the prompt plus any documents, conversation history and tool results sent with it; every token in it is billed at the input rate.

What inputs and outputs does GLM-5V-Turbo support?

GLM-5V-Turbo accepts text, images and video as input and produces text. Its listed capabilities are function calling.

How does GLM-5V-Turbo compare with GLM-5-Turbo?

GLM-5V-Turbo and GLM-5-Turbo both cost $1.20 per 1M input tokens ($4.00 vs $4.00 per 1M output tokens). Both have a 205K-token context window.

What are cheaper alternatives to GLM-5V-Turbo?

Models from other creators with a lower input rate and at least GLM-5V-Turbo's 205K-token context window: Qwen3-Coder-Plus at $1.00 per 1M input tokens (1M context), Grok Build 0.1 at $1.00 per 1M input tokens (256K context) and Qwen3-VL-235B-A22B Thinking at $0.98 per 1M input tokens (262K context). Rates are the cheapest listed provider for each; whether the quality holds for a given task is a separate question.

Cheaper alternatives to GLM-5V-Turbo