Alibaba Cloud logo Qwen3.8-27B

View website

Released August 2026 under Apache 2.0, a 27.8B dense model using a hybrid decoder that alternates Gated DeltaNet linear-attention layers with grouped-query full attention. Thinking mode is on by default and can be toggled per request with tunable reasoning effort; Alibaba reports it ahead of the larger Qwen3.7-Plus on coding and office tasks.

At a glance

Context window
262K tokens
Max output
131K tokens
Modalities
Text Text Image Image Video Video Text Text

Capabilities

Function calling

Function calling

Connect to external tools, APIs, and systems.

Structured output

Structured output

Return responses in structured formats like JSON.

Pricing by provider

Provider Input / 1M tokens Output / 1M tokens
Alibaba Cloud logo Alibaba Cloud Self-hosted Self-hosted View

Heads up: We do our best to keep these specs & prices accurate. However, cloud costs may fluctuate based on region, usage, and other factors not listed here. These are estimates based on common setups and are for informational purposes only. Always verify current rates & exact specs with the provider before provisioning.

Compare with other models

Estimated prices shown. Actual costs may vary based on context length, batch size, caching, and provider-specific pricing tiers.