AMD MI355X

AMD MI355X

AMD's next-gen CDNA 4 competitor to Nvidia Blackwell.

Last update 11 minutes ago
Compare vs other GPUs →
Aggregating historical prices...

Weekly median price per GPU per hour. AMD MI355X price history (JSON).

Key Specifications

Architecture
CDNA 4
Memory per GPU
288 GB HBM3e
Memory bandwidth
8,000 GB/s
GPU interconnect
Infinity Fabric 1,075 GB/s
Release date
Q2 2025

MI355X Pricing and Availability

Of 4 listings across 4 cloud providers, none are verified in stock. On-demand is listed from $8.60 per GPU per hour and spot from $4.50.

4 providers

How this list works

The two views

By provider shows one card per company, placed where its best offer ranked. By configuration lists every offer.

Order

What you can rent comes first: in stock, then waitlist, then not reported, then out of stock. Priced offers rank ahead of quote-only, and on-demand ahead of other billing types.

Within a group, five factors set the order:

  1. Location: datacenter proximity, blended with provider HQ.
  2. Price: hourly price, per GPU and in total.
  3. Billing type: reserved ahead of spot, spot ahead of quote-only.
  4. Specs: more VRAM, vCPUs and RAM.
  5. Provider diversity: a provider's repeat rows rank slightly lower, so one company can't take all the top spots.

Search

We match the provider, its country, GPU model, form factor, billing type, availability and instance name. Matching is partial and case-insensitive. Everyday words work too, so "interruptible" finds spot and "sold out" finds out of stock.

Transparency and funding

  • Ads and sponsors: Paid placements sit at the top of the list and are always labeled as sponsored content. Sponsorship never influences the organic ranking itself.
  • Affiliates: Affiliate links are marked. We may earn a commission if you click them, but commissions never affect the order.
  • Prices: Shown in USD, converted at daily reference rates where a provider publishes in another currency. A month means 720 hours. How we estimate costs.
Oracle Cloud logo

United States of America flag USA 1 config 8x

On-Demand from $8.60
From $8.60 / GPU / hr On-Demand Visit website
DigitalOcean logo

United States of America flag USA 1 config 8x

Spot from $4.50
From $4.50 / GPU / hr Spot Visit website
TensorWave logo

United States of America flag USA 1 config 8x

Custom on request
Pricing On request Custom Visit website
Vultr logo

Vultr

Out of stock

United States of America flag USA 1 config 8x

Spot from $2.59
From $2.59 / GPU / hr Spot Visit website

Heads up: Prices are estimates from published rates for common setups and vary by region and usage; verify with the provider before provisioning.

What size AI models can the MI355X run?

One MI355X has 288 GB of VRAM. In practice, that's enough memory for roughly 460B parameters at 4-bit or 125B at 16-bit, assuming a 32K context. Below are some open-weight LLMs, with the estimated memory and GPUs each one needs.

Model Parameters 4-bitINT4 / FP4 8-bitFP8 / INT8 16-bitFP16 / BF16
Alibaba Cloud logo Qwen3.8-27B
27B
8 GPUs 18 GB VRAM
8 GPUs 30 GB VRAM
8 GPUs 58 GB VRAM
Google Cloud logo Gemma 4 31B
30.7B
8 GPUs 22 GB VRAM
8 GPUs 35 GB VRAM
8 GPUs 69 GB VRAM
OpenAI logo GPT-OSS-120B
117B 5.1B active
8 GPUs 67 GB VRAM
8 GPUs 120 GB VRAM
8 GPUs 237 GB VRAM
DeepSeek logo DeepSeek V4 Flash
284B 13B active
8 GPUs 160 GB VRAM
8 GPUs 287 GB VRAM
8 GPUs 573 GB VRAM
Z.AI logo GLM-5.3-Flash
320B 18B active
8 GPUs 178 GB VRAM
8 GPUs 322 GB VRAM
8 GPUs 642 GB VRAM
MiniMax logo MiniMax-M3
428B 23B active
8 GPUs 239 GB VRAM
8 GPUs 432 GB VRAM
8 GPUs 862 GB VRAM
Moonshot AI logo Kimi K3
2.8T 104B active
8 GPUs 1,544 GB VRAM
11 GPUs 2,804 GB VRAM
22 GPUs 5,605 GB VRAM

Estimates based on the median on-demand rate. Memory is weights plus FP8 KV cache at 32K context per request (FP16/BF16 in the 16-bit column). GPU counts assume 90% of advertised VRAM is usable. No guarantee of runtime support, usable performance, or that a matching quantized build exists. How we estimate costs.

Technical Specifications

AMD MI355X · Per GPU

Compute · dense
FP4 10,066.3 TFLOPS
FP8 5,033.2 TFLOPS 10,066.4 with sparsity
FP16 / BF16 2,516.6 TFLOPS 5,033.2 with sparsity
INT8 5,033.2 TOPS 10,066.4 with sparsity
FP32 157.3 TFLOPS
FP64 78.6 TFLOPS
Precision support FP4FP6FP8FP16BF16FP32FP64INT8
Memory
Capacity 288 GB HBM3e
Bandwidth 8,000 GB/s
Bus width 8,192-bit
ECC Yes
Silicon
Architecture CDNA 4
Process TSMC 3nm / 6nm FinFET
Transistors 185 billion
Shader cores 16,384 Stream processors
Matrix cores 1,024
Compute units 256
Fabric and host
GPU interconnect Infinity Fabric 1,075 GB/s
Host interface PCIe 5.0 x16
Power
Board power 1,400 W
Platform
Partitioning Up to 8 partitions

Source: official AMD MI355X datasheet.

Frequently Asked Questions

Why choose the MI355X?

288GB HBM3e with FP4 support on CDNA 4 architecture. Next-generation AMD compute for both training and inference. Competitive with Nvidia Blackwell on memory capacity.

When is the MI355X not a good fit?

Latest generation with minimal cloud availability. ROCm software stack, while rapidly improving, may lag behind CUDA for some workloads. Evaluate software compatibility before committing.

Which cloud providers offer the MI355X?

The MI355X is listed by 4 providers.

What data types does the MI355X support?

The MI355X supports 8 precision formats. Training: BF16, FP16, FP32. Inference: FP4, FP6, FP8, INT8. Scientific: FP64.

Does the MI355X support Infinity Fabric?

Yes. 1,075 GB/s bidirectional.

How much does the MI355X cost per hour?

As of August 30, 2026, we track 4 listings from 4 providers. Prices are per GPU per hour.

Billing type Listings Median / GPU / hr Cheapest in stock
On-demand 1 $8.60
Spot 2 $4.50
Custom contract 1 On request

No median for on-demand and spot: we only show this when at least 3 providers list the GPU on that billing type.

Alternatives to AMD MI355X

Last updated