We match your deposits, up to $150 in free credits. Ends October 1.Learn more  

gpu catalog: l40s

Rent L40S from $0.610/hr

The universal data center Ada card: 48 GB and strong FP8 throughput make it a favorite for mixed inference, fine-tuning, and rendering fleets.

Live platform rates

25 GPU types online
GPUVRAMRegionsAvailabilityPrice / GPU
L40S48G4 regionshigh availability$0.80/hrlaunch

The universal data center Ada card: 48 GB and strong FP8 throughput make it a favorite for mixed inference, fine-tuning, and rendering fleets.

VRAM48 GB GDDR6
Memory bandwidth0.9 TB/s
FP16 tensor183 TFLOPS
ArchitectureAda Lovelace
InterconnectPCIe 4.0
Board power350 W
CUDA cores18,176
Released2023
Regionsasia-pacific, eu-west, us-central, us-east
L40Scommunity48G5 regionsavailable$0.61/hrlaunch
RTX PRO 600096G5 regionshigh availability$1.99/hrlaunch
L4048G2 regionshigh availability$0.78/hrlaunch
RTX A600048G2 regionshigh availability$0.40/hrlaunch
RTX 4090community24G7 regionshigh availability$0.31/hrlaunch
A10080G5 regionshigh availability$1.07/hrlaunch
RTX 5090community32G7 regionshigh availability$0.35/hrlaunch
RTX 6000 ADA48G2 regionshigh availability$0.65/hrlaunch
RTX PRO 6000community96G4 regionshigh availability$1.69/hrlaunch
A100community80G6 regionshigh availability$0.95/hrlaunch
H200 SXMcommunity141G5 regionshigh availability$3.59/hrlaunch
RTX PRO 450032G1 regionhigh availability$0.72/hrlaunch
H200 SXM141G6 regionshigh availability$3.99/hrlaunch
H100 SXM80G5 regionshigh availability$2.54/hrlaunch
RTX 3090community24G6 regionshigh availability$0.15/hrlaunch
RTX 409024G4 regionshigh availability$0.74/hrlaunch
RTX A4000community16G5 regionsavailable$0.08/hrlaunch
L424G4 regionsavailable$0.44/hrlaunch
H100 NVLcommunity94G5 regionsavailable$2.37/hrlaunch
H200 NVLcommunity141G5 regionsavailable$3.47/hrlaunch
A100 40GBcommunity40G5 regionsavailable$0.67/hrlaunch
RTX 3080community10G3 regionsavailable$0.07/hrlaunch
H100 PCIE80G2 regionsavailable$1.98/hrlaunch
H100 NVL94G3 regionsavailable$3.19/hrlaunch
H100 SXMcommunity80G3 regionsavailable$1.87/hrlaunch
H100 PCIEcommunity80G2 regionsavailable$2.54/hrlaunch
B200community192G3 regionsavailable$5.32/hrlaunch
RTX 5080community16G5 regionsrunning low$0.16/hrlaunch
RTX 4080community16G4 regionsrunning low$0.17/hrlaunch
A4048G2 regionsrunning low$0.49/hrlaunch
RTX 509032G2 regionsrunning low$0.99/hrlaunch
B200192G1 regionrunning low$6.79/hrlaunch
B300288G2 regionsrunning low$7.89/hrlaunch
RTX A5000community24G4 regionsrunning low$0.15/hrlaunch
L4community24G2 regionsrunning low$0.27/hrlaunch
RTX A6000community48G3 regionsrunning low$0.41/hrlaunch
RTX 309024G1 regionrunning low$0.50/hrlaunch
H200 NVL141G1 regionrunning low$3.29/hrlaunch
A100 40GB40G1 regionrunning low$0.89/hrlaunch

lowest live on-demand rate per GPU, refreshed continuously · per-second billing

L40S price comparison (September 2026)

ProviderPrice / GPU / hrNotes
GPU.ai (Secure Cloud)$0.800/hrVetted data center capacity, per-second billing
GPU.ai (Community Cloud)$0.610/hrMarketplace capacity from independent suppliers
CoreWeave$2.25/hrPublic on-demand list rate
AWS$1.86/hrPublic on-demand list rate
Azure$2.50/hrPublic on-demand list rate

competitor rates are public on-demand list prices, refreshed daily · gpu.ai prices are live floor rates with zero markup · market data via Compute Prices

Compare and explore

faq

L40S rental questions

How much does it cost to rent an L40S in the cloud?

L40S cloud instances on GPU.ai currently start at $0.610 per GPU per hour, billed per second. That is the live floor price across every provider in our fleet. We route your launch to the cheapest provider with stock and add zero markup.

How much VRAM does the L40S have?

The NVIDIA L40S has 48 GB of GDDR6 memory with up to 864 GB/s of bandwidth. GPU.ai reports the minimum VRAM actually delivered across our fleet, so the figure you see at launch is the figure you get.

Can I rent a multi-GPU L40S node?

Yes: real 8× L40S nodes are launchable right now, from $0.800 per GPU per hour. We only advertise node sizes that exist in live inventory, not theoretical configurations.

How does GPU.ai's L40S pricing compare to AWS, Azure, and CoreWeave?

GPU.ai aggregates a dozen-plus GPU clouds and passes through each provider's raw price with no markup, so the same silicon is typically far cheaper than hyperscaler list rates. The comparison table on this page shows the current public on-demand list prices side by side.

How does billing work?

Billing is per second, starting when your instance is SSH-ready and stopping the moment you terminate. There are no minimum commitments, no reservation fees, and no egress surprises: you pay the listed hourly rate pro-rated to the second.