Cohere logo

Command A+

Open weights · Apache 2.0

Cohere's first mixture-of-experts model, released May 2026 with 218B total parameters (25B active), combining vision, agentic work, optional reasoning and translation across 48 languages. Apache 2.0.

Pricing Open weights

No provider we track hosts it. Self-hosting starts at $2.90 an hour: see the estimate.

Cohere 1 provider

Key Specifications

Context window
128K tokens
Max output
64K tokens
Inputs
Text, Image Outputs: Text

Command A+ pricing by provider

Provider Input / 1M tokens Output / 1M tokens Cost at 10M in + 2M out
Cohere logo Cohere Creator Open weights, no hosted price View

Heads up: We do our best to keep these specs & prices accurate. However, cloud costs may fluctuate based on region, usage, and other factors not listed here. These are estimates based on common setups and are for informational purposes only. Always verify current rates & exact specs with the provider before provisioning. LLM rates are base-tier, on-demand prices per 1M tokens; cached-input, batch and long-context tiers are not included.

Compare every model at this volume in the LLM cost calculator.

Capabilities

Function calling

Function calling

Connect to external tools, APIs, and systems.

Structured output

Structured output

Return responses in structured formats like JSON.

Estimated cost to self-host Command A+

Command A+ has 218B parameters, about 25B active per token. At 4-bit it needs about 131 GB of GPU memory, at 8-bit 262 GB and at BF16 523 GB, counting 20% on top of the weights for KV cache and runtime overhead.

Precision Memory needed Cheapest rentals that fit Per month, 24/7
4-bit 131 GB Amd logo MI300X ($2.90/hr) Amd logo MI325X ($3.07/hr) $2,088
8-bit 262 GB Nvidia logo B300 ($7.89/hr, tight) $5,681
BF16 523 GB Amd logo 4× MI300X ($11.60/hr) Amd logo 4× MI325X ($12.28/hr) $8,352

Memory is parameters × bytes per weight at each precision, plus 20% for KV cache and runtime overhead. Rentals are the cheapest cards that hold it, at provider-weighted median on-demand prices for the week of August 17, 2026, consumer cards included, in nodes of up to 8 GPUs; a fit with under 15% headroom is marked tight. A month is 720 hours. See the best-value GPUs guide for the same table across models, and cloud GPU pricing for every card.

More from Cohere

Model Context Input / 1M Output / 1M
Cohere logo Command A Reasoning Open weights 256K No hosted price
Cohere logo Command A Vision Open weights 128K No hosted price
Cohere logo Command R 08-2024 Open weights 128K $0.15 $0.60
Cohere logo Command A Translate Open weights 8K No hosted price
Cohere logo Command R+ 08-2024 Open weights 128K $2.50 $10.00
Cohere logo Command A Open weights 256K $2.50 $10.00
Cohere logo Command R7B Open weights 128K $0.04 $0.15

Models near this price

The models nearest this one by input rate, five each way, at each one's cheapest listed provider.

Frequently Asked Questions

How much does Command A+ cost?

Cohere publishes Command A+'s weights and lists no hosted price. Running it yourself costs whatever the hardware does; see cloud GPU pricing.

Which providers offer Command A+?

One provider lists Command A+: Cohere (open weights, no hosted price).

What is Command A+'s context window?

Command A+ accepts up to 128K tokens of input per request and returns up to 64K tokens per response. The context window is the prompt plus any documents, conversation history and tool results sent with it; every token in it is billed at the input rate.

What inputs and outputs does Command A+ support?

Command A+ accepts text and images as input and produces text. Its listed capabilities are function calling and structured output.

Can I self-host Command A+?

Yes. Cohere publishes Command A+'s weights under the Apache 2.0 license. At 4-bit it needs about 131 GB of GPU memory; the cheapest rental that holds it is MI300X at $2.90 per hour, about $2,088 a month. The estimate above prices 8-bit and BF16 too.