Nvidia L40

Nvidia L40

Data center GPU for combined AI inference and visualization.

Last update 8 minutes ago
Compare vs other GPUs →
Aggregating historical prices...

Key Specifications

Architecture
Ada Lovelace
Memory
48GB GDDR6
Memory Bandwidth
864 GB/s
Release date
Q4 2022

L40 Pricing and Availability

Listings for the L40 reach $1.64/hr, often reflecting a premium for high availability. However, you might be able to find available instances from as low as $0.58/hr per GPU (on-demand). Spot instances start lower, at $0.47/hr per GPU.

12 providers

How this list works

The two views

By provider shows one card per company, placed where its best offer ranked. By configuration lists every offer.

Order

What you can rent comes first: in stock, then waitlist, then not reported, then out of stock. Priced offers rank ahead of quote-only, and on-demand ahead of other billing types.

Within a group, five factors set the order, largest first:

  1. Location: datacenter proximity, blended with provider HQ.
  2. Price: hourly price, per GPU and in total.
  3. Billing type: reserved ahead of spot, spot ahead of quote-only.
  4. Specs: more VRAM, vCPUs and RAM.
  5. Provider diversity: in By configuration, a provider's repeat rows rank slightly lower, so one company can't take all the top spots.

Search

We match the provider, GPU model, form factor, billing type, availability and instance name. Matching is partial and case-insensitive. Everyday words work too, so "interruptible" finds spot and "sold out" finds out of stock.

Transparency and funding

  • Ads and sponsors: Paid placements sit at the top of the list and are always labeled as sponsored content. Sponsorship never influences the organic ranking itself.
  • Affiliates: Affiliate links are marked. We may earn a commission if you click them, but commissions never affect the order.
  • Prices: Shown in USD, converted at daily reference rates where a provider publishes in another currency. A month means 720 hours. How we estimate costs.
Runpod logo

Runpod Our sponsor

USA 4 configs 1x-8x

On-Demand from $0.69 Reserved on request
From $0.69 / GPU / hr On-Demand Visit website
Massed Compute logo

Massed Compute In stock

USA 6 configs 1x-4x

On-Demand from $0.86 Spot from $0.77
From $0.86 / GPU / hr On-Demand Visit website
Runcrate logo

Runcrate In stock

USA 1 config 1x

On-Demand from $0.97
From $0.97 / GPU / hr On-Demand PCIe Visit website
Vast.ai logo

Vast.ai In stock

USA 3 configs 1x

On-Demand from $0.58 Reserved from $0.58 Spot from $0.47
From $0.58 / GPU / hr On-Demand Visit website
Thunder Compute logo

Thunder Compute In stock

USA 4 configs 1x-8x

On-Demand from $0.79
From $0.79 / GPU / hr On-Demand Visit website
Sesterce logo

Sesterce In stock

France 7 configs 1x-8x

On-Demand from $0.97
From $0.97 / GPU / hr On-Demand Visit website
GPU.ai logo

GPU.ai In stock

UAE 4 configs 1x-8x

On-Demand from $0.78
From $0.78 / GPU / hr On-Demand Visit website
Spheron logo

Spheron In stock

Singapore 9 configs 1x-8x

On-Demand from $1.06 Spot from $0.88
From $1.06 / GPU / hr On-Demand PCIe Visit website
TensorDock logo

TensorDock

USA 1 config 1x

On-Demand from $1.06
From $1.06 / GPU / hr On-Demand Visit website
Hyperstack logo

Hyperstack

UK 3 configs 1x

On-Demand from $1.00 Reserved from $0.70 Spot from $0.80
From $1.00 / GPU / hr On-Demand Visit website
Oblivus logo

Oblivus

UK 1 config 1x

On-Demand from $1.05
From $1.05 / GPU / hr On-Demand Visit website
CoreWeave logo

CoreWeave

USA 2 configs 8x

On-Demand from $1.25 Spot from $0.78
From $1.25 / GPU / hr On-Demand Visit website

Heads up: We do our best to keep these specs & prices accurate. However, cloud costs may fluctuate based on region, usage, and other factors not listed here. These are estimates based on common setups and are for informational purposes only. Always verify current rates & exact specs with the provider before provisioning.

Frequently Asked Questions

Why choose the L40?

48GB GDDR6 with Ada Lovelace architecture. Good for visualization, rendering, and inference workloads. FP8 Tensor Core support for efficient AI inference.

When is the L40 not a good fit?

PCIe-only. Lower inference throughput than L40S for pure AI workloads. If you don't need the visualization features, the L40S is typically a better value.

Are L40 prices going up or down?

The median on-demand price across providers has fallen about 8% since August 2025, from $1.09 to $1.00/hr per GPU. This reflects the market-wide median, which moves both when providers change prices and when lower or higher priced offerings enter the market.

What size AI models can the L40 run?

With 48GB of VRAM, the L40 can typically run models up to about 30B parameters in FP16, or 70B-class models in 4-bit quantized form for inference.

How much VRAM does the L40 have?

The L40 has 48GB of VRAM. Multi-GPU setups increase total memory, but that memory is not automatically pooled across GPUs.

What is the L40's memory bandwidth?

The L40 has 864 GB/s of memory bandwidth. Higher bandwidth helps with faster data transfer between GPU memory and compute cores.

What data types does the L40 support?

The L40 supports 7 precision formats. Training: BF16, FP16, TF32, FP32. Inference: FP8, INT4, INT8.

Does the L40 support NVLink?

No. The L40 is a PCIe-only GPU with no NVLink, so it is better suited to single-GPU inference and smaller-scale workloads than large distributed training jobs.

How much does the L40 cost per hour?

L40 pricing currently ranges from $0.47/hr to $1.64/hr per GPU, depending on the provider, instance type, and billing model.

How much does the L40 cost per month?

At 720 hours per month, one L40 can cost between $339.19 to $1,180.08 per month, depending on the provider. Reserved and spot pricing can lower that further.

Which cloud providers offer the L40?

The L40 is available from 12 cloud providers, including Spheron, Sesterce, Runpod. Pricing and availability vary by region and billing model.

Can I rent the L40 in the cloud?

Yes. We currently track 45 L40 listings across 12 cloud providers:

Billing type Listings Avg $/GPU/hr
On-demand 33 $0.98/hr
Reserved 3 $0.64/hr
Spot 9 $0.78/hr

Technical Specifications

GPU Architecture NVIDIA Ada Lovelace architecture
GPU Memory 48GB GDDR6
Memory Bandwidth 864GB/s
Interconnect Interface PCIe Gen4 x16 (64GB/s bi-directional)
CUDA Cores 18,176
Third-Generation RT Cores 142
Fourth-Generation Tensor Cores 568
RT Core Performance TFLOPS 209
FP32 TFLOPS 90.5
TF32 Tensor Core TFLOPS 90.5
BFLOAT16 Tensor Core TFLOPS 181.05
FP16 Tensor Core TFLOPS 181.05
FP8 Tensor Core TFLOPS 362
Peak INT8 Tensor TOPS 362
Peak INT4 Tensor TOPS 724
Form Factor 4.4" (H) x 10.5" (L) - dual slot
Display Ports 4x DisplayPort 1.4a
Max Power Consumption 300W
NVLink Support No

Source: official Nvidia L40 datasheet.

Alternatives to Nvidia L40

Last updated