Nvidia L40S

Nvidia L40S

Cost-effective data center GPU for AI inference and media workloads.

Last update 1 minute ago
Compare vs other GPUs →
Aggregating historical prices...

Key Specifications

Architecture
Ada Lovelace
Memory
48GB GDDR6
Memory Bandwidth
864 GB/s
Release date
Q3 2023

L40S Pricing and Availability

Listings for the L40S reach $7.58/hr, often reflecting a premium for high availability. However, you might be able to find available instances from as low as $0.47/hr per GPU (on-demand). Spot instances start lower, at $0.33/hr per GPU.

32 providers

How this list works

The two views

By provider shows one card per company, placed where its best offer ranked. By configuration lists every offer.

Order

What you can rent comes first: in stock, then waitlist, then not reported, then out of stock. Priced offers rank ahead of quote-only, and on-demand ahead of other billing types.

Within a group, five factors set the order, largest first:

  1. Location: datacenter proximity, blended with provider HQ.
  2. Price: hourly price, per GPU and in total.
  3. Billing type: reserved ahead of spot, spot ahead of quote-only.
  4. Specs: more VRAM, vCPUs and RAM.
  5. Provider diversity: in By configuration, a provider's repeat rows rank slightly lower, so one company can't take all the top spots.

Search

We match the provider, GPU model, form factor, billing type, availability and instance name. Matching is partial and case-insensitive. Everyday words work too, so "interruptible" finds spot and "sold out" finds out of stock.

Transparency and funding

  • Ads and sponsors: Paid placements sit at the top of the list and are always labeled as sponsored content. Sponsorship never influences the organic ranking itself.
  • Affiliates: Affiliate links are marked. We may earn a commission if you click them, but commissions never affect the order.
  • Prices: Shown in USD, converted at daily reference rates where a provider publishes in another currency. A month means 720 hours. How we estimate costs.
Lyceum logo

Lyceum Our sponsor

Germany 5 configs 1x-8x

On-Demand from $1.19 Reserved on request
From $1.19 / GPU / hr On-Demand Visit website
Packet·ai logo

Packet·ai In stock

USA 1 config 1x

On-Demand from $0.92
From $0.92 / GPU / hr On-Demand Visit website
Novita logo

Novita In stock

USA 8 configs 1x-8x

On-Demand from $0.55 Spot from $0.28
From $0.55 / GPU / hr On-Demand Visit website
Runpod logo

Runpod In stock

USA 7 configs 1x-8x

On-Demand from $0.79 Reserved on request
From $0.79 / GPU / hr On-Demand Visit website
Massed Compute logo

Massed Compute In stock

USA 4 configs 1x-8x

On-Demand from $0.88
From $0.88 / GPU / hr On-Demand Visit website
Runcrate logo

Runcrate In stock

USA 1 config 1x

On-Demand from $0.97
From $0.97 / GPU / hr On-Demand PCIe Visit website
Crusoe logo

Crusoe In stock

USA 7 configs 1x-10x

On-Demand from $1.50 Reserved on request Spot on request
From $1.50 / GPU / hr On-Demand Visit website
Amazon Web Services logo

AWS In stock

USA 24 configs 1x-8x

On-Demand from $1.86 Reserved from $0.80
From $1.86 / GPU / hr On-Demand Visit website
Vast.ai logo

Vast.ai In stock

USA 9 configs 1x-4x

On-Demand from $0.47 Reserved from $0.73 Spot from $0.33
From $0.47 / GPU / hr On-Demand Visit website
Sesterce logo

Sesterce In stock

France 7 configs 1x-4x

On-Demand from $0.97
From $0.97 / GPU / hr On-Demand Visit website
Koyeb logo

Koyeb In stock

France 1 config 1x

On-Demand from $1.20
From $1.20 / GPU / hr On-Demand Visit website
UpCloud logo

UpCloud In stock

Finland 22 configs 1x-3x

On-Demand from $1.19 Spot from $0.89
From $1.19 / GPU / hr On-Demand Visit website
Scaleway logo

Scaleway In stock

France 4 configs 1x-8x

On-Demand from $1.72
From $1.72 / GPU / hr On-Demand Visit website
GPU.ai logo

GPU.ai In stock

UAE 3 configs 1x-4x

On-Demand from $0.47
From $0.47 / GPU / hr On-Demand Visit website
Spheron logo

Spheron In stock

Singapore 24 configs 1x-4x

On-Demand from $2.11 Spot from $1.07
From $2.11 / GPU / hr On-Demand PCIe Visit website
Verda logo

Verda In stock

Finland 28 configs 1x-8x

On-Demand from $1.37 Reserved from $1.03 Spot from $0.69
From $0.69 / GPU / hr Spot Visit website
Fly.io logo

Fly.io

USA 1 config 1x

On-Demand from $1.25
From $1.25 / GPU / hr On-Demand Visit website
Geodd logo

Geodd

USA 4 configs 1x-8x

On-Demand from $1.00
From $1.00 / GPU / hr On-Demand PCIe Visit website
DigitalOcean logo

DigitalOcean

USA 1 config 1x

On-Demand from $1.57
From $1.57 / GPU / hr On-Demand Visit website
Cerebrium logo

Cerebrium

USA 1 config 1x

On-Demand from $1.95
From $1.95 / GPU / hr On-Demand Visit website
Replicate logo

Replicate

USA 2 configs 1x-2x

On-Demand from $3.51
From $3.51 / GPU / hr On-Demand Visit website
Civo logo

Civo

UK 20 configs 1x-8x

On-Demand from $1.29 Reserved from $0.89
From $1.29 / GPU / hr On-Demand Visit website
Nebius logo

Nebius

Netherlands 4 configs 1x

On-Demand from $1.55 Spot from $0.74
From $1.55 / GPU / hr On-Demand PCIe Visit website
OVHcloud logo

OVH

France 3 configs 1x-4x

On-Demand from $1.80
From $1.80 / GPU / hr On-Demand Visit website
Gcore logo

Gcore

Luxembourg 1 config 8x

On-Demand from $1.27
From $1.27 / GPU / hr On-Demand PCIe Visit website
Elastx logo

Elastx

Sweden 1 config 1x

On-Demand from $2.15
From $2.15 / GPU / hr On-Demand Visit website
CoreWeave logo

CoreWeave

USA 2 configs 8x

On-Demand from $2.25 Spot from $0.98
From $2.25 / GPU / hr On-Demand Visit website
Oracle Cloud logo

Oracle Cloud

USA 1 config 4x

On-Demand from $3.50
From $3.50 / GPU / hr On-Demand Visit website
AceCloud logo

AceCloud

India 4 configs 1x

On-Demand from $1.87 Reserved from $1.37
From $1.87 / GPU / hr On-Demand Visit website
Cyfuture AI logo

Cyfuture AI

India 16 configs 1x-8x

On-Demand from $1.25 Reserved from $0.60
From $1.25 / GPU / hr On-Demand Visit website
Contabo logo

Contabo

Germany 2 configs 1x

Reserved from $1.33
From $1.33 / GPU / hr Reserved (1mo) Visit website
Vultr logo

Vultr Out of stock

USA 4 configs 1x-8x

On-Demand from $1.67 Spot from $1.49
From $1.67 / GPU / hr On-Demand Visit website

Heads up: We do our best to keep these specs & prices accurate. However, cloud costs may fluctuate based on region, usage, and other factors not listed here. These are estimates based on common setups and are for informational purposes only. Always verify current rates & exact specs with the provider before provisioning.

Frequently Asked Questions

Why choose the L40S?

48GB GDDR6 with Ada Lovelace FP8 support. Strong inference throughput at a lower cost than A100 or H100. AV1 hardware encoding makes it versatile for media and AI workloads.

When is the L40S not a good fit?

PCIe-only with no NVLink, limiting multi-GPU training performance. GDDR6 memory bandwidth (864 GB/s) is significantly lower than HBM-based data center GPUs.

Are L40S prices going up or down?

The median on-demand price has held steady around $1.45/hr per GPU across providers.

What size AI models can the L40S run?

With 48GB of VRAM, the L40S can typically run models up to about 30B parameters in FP16, or 70B-class models in 4-bit quantized form for inference.

How much VRAM does the L40S have?

The L40S has 48GB of VRAM. Multi-GPU setups increase total memory, but that memory is not automatically pooled across GPUs.

What is the L40S's memory bandwidth?

The L40S has 864 GB/s of memory bandwidth. Higher bandwidth helps with faster data transfer between GPU memory and compute cores.

What data types does the L40S support?

The L40S supports 7 precision formats. Training: BF16, FP16, TF32, FP32. Inference: FP8, INT4, INT8.

Does the L40S support NVLink?

No. The L40S is a PCIe-only GPU with no NVLink, so it is better suited to single-GPU inference and smaller-scale workloads than large distributed training jobs.

How much does the L40S cost per hour?

L40S pricing currently ranges from $0.33/hr to $7.58/hr per GPU, depending on the provider, instance type, and billing model.

How much does the L40S cost per month?

At 720 hours per month, one L40S can cost between $240.26 to $5,455.58 per month, depending on the provider. Reserved and spot pricing can lower that further.

Which cloud providers offer the L40S?

The L40S is available from 32 cloud providers, including Verda, Amazon Web Services, Spheron. Pricing and availability vary by region and billing model.

Can I rent the L40S in the cloud?

Yes. We currently track 222 L40S listings across 32 cloud providers:

Billing type Listings Avg $/GPU/hr
On-demand 108 $1.74/hr
Reserved 75 $1.31/hr
Spot 39 $1.11/hr

Technical Specifications

GPU Architecture NVIDIA Ada Lovelace Architecture
GPU Memory 48GB GDDR6
Memory Bandwidth 864GB/s
Interconnect Interface PCIe Gen4 x16: 64GB/s bidirectional
NVIDIA Ada Lovelace Architecture-Based CUDA® Cores 18,176
NVIDIA Third-Generation RT Cores 142
NVIDIA Fourth-Generation Tensor Cores 568
RT Core Performance TFLOPS 209
FP32 TFLOPS 91.6
TF32 Tensor Core TFLOPS 183
BFLOAT16 Tensor Core TFLOPS 362.05
FP16 Tensor Core 362.05
FP8 Tensor Core 733
Peak INT8 Tensor TOPS 733
Peak INT4 Tensor TOPS 733
Form Factor 4.4" (H) x 10.5" (L), dual slot
Display Ports 4x DisplayPort 1.4a
Max Power Consumption 350W
Power Connector 16-pin
Thermal Passive
Virtual GPU (vGPU) Software Support Yes
vGPU Profiles Supported See the virtual GPU licensing guide
NVENC, NVDEC 3x, 3x (includes AV1 encode and decode)
Secure Boot With Root of Trust Yes
NEBS Ready Level 3
MIG Support No
NVIDIA® NVLink® Support No

Source: official Nvidia L40S datasheet.

Alternatives to Nvidia L40S

Last updated