We match your deposits, up to $150 in free credits. Ends October 1.Learn more  

gpu catalog: a100 40gb

Rent A100 40GB from $0.670/hr

The original A100: same compute as the 80 GB part with half the memory, often the cheapest NVLink-capable data center silicon on the market.

Live platform rates

25 GPU types online
GPUVRAMRegionsAvailabilityPrice / GPU
A100 40GBcommunity40G5 regionsavailable$0.67/hrlaunch

The original A100: same compute as the 80 GB part with half the memory, often the cheapest NVLink-capable data center silicon on the market.

VRAM40 GB HBM2
Memory bandwidth1.6 TB/s
FP16 tensor312 TFLOPS
ArchitectureAmpere
InterconnectNVLink 3 (600 GB/s)
Board power400 W
CUDA cores6,912
Released2020
Regionsasia-pacific, eu-east, eu-west, us-east, us-west
L40S48G3 regionshigh availability$0.80/hrlaunch
L4048G2 regionshigh availability$0.78/hrlaunch
RTX A600048G1 regionhigh availability$0.45/hrlaunch
RTX PRO 600096G4 regionshigh availability$1.99/hrlaunch
A10080G5 regionshigh availability$1.07/hrlaunch
RTX 4090community24G7 regionshigh availability$0.31/hrlaunch
RTX 5090community32G7 regionshigh availability$0.37/hrlaunch
RTX 6000 ADA48G1 regionhigh availability$0.65/hrlaunch
A100community80G7 regionshigh availability$0.95/hrlaunch
RTX PRO 450032G2 regionshigh availability$0.72/hrlaunch
H200 SXMcommunity141G5 regionshigh availability$3.59/hrlaunch
H100 SXM80G5 regionshigh availability$2.54/hrlaunch
H200 SXM141G6 regionshigh availability$3.99/hrlaunch
RTX 3090community24G7 regionshigh availability$0.15/hrlaunch
RTX A4000community16G5 regionsavailable$0.08/hrlaunch
L424G4 regionsavailable$0.44/hrlaunch
RTX 409024G3 regionsavailable$0.74/hrlaunch
L40Scommunity48G4 regionsavailable$0.61/hrlaunch
H100 NVLcommunity94G4 regionsavailable$2.59/hrlaunch
H100 SXMcommunity80G4 regionsavailable$1.87/hrlaunch
H200 NVLcommunity141G5 regionsavailable$3.47/hrlaunch
RTX 3080community10G3 regionsavailable$0.07/hrlaunch
H100 NVL94G3 regionsavailable$3.19/hrlaunch
B200192G2 regionsavailable$6.79/hrlaunch
B300288G2 regionsavailable$7.89/hrlaunch
H100 PCIE80G2 regionsavailable$1.98/hrlaunch
H100 PCIEcommunity80G2 regionsavailable$2.48/hrlaunch
RTX 5080community16G6 regionsrunning low$0.17/hrlaunch
RTX 4080community16G4 regionsrunning low$0.17/hrlaunch
RTX A5000community24G2 regionsrunning low$0.23/hrlaunch
B200community192G2 regionsrunning low$5.32/hrlaunch
L4community24G2 regionsrunning low$0.27/hrlaunch
RTX 509032G1 regionrunning low$0.99/hrlaunch
RTX PRO 6000community96G1 regionrunning low$1.69/hrlaunch
RTX A6000community48G2 regionsrunning low$0.42/hrlaunch
H200 NVL141G1 regionrunning low$3.29/hrlaunch
V100community16G1 regionrunning low$0.17/hrlaunch

lowest live on-demand rate per GPU, refreshed continuously · per-second billing

A100 40GB price comparison (September 2026)

ProviderPrice / GPU / hrNotes
GPU.ai (Secure Cloud)Vetted data center capacity, per-second billing
GPU.ai (Community Cloud)$0.670/hrMarketplace capacity from independent suppliers
CoreWeavePublic on-demand list rate
AWSPublic on-demand list rate
AzurePublic on-demand list rate

competitor rates are public on-demand list prices, refreshed daily · gpu.ai prices are live floor rates with zero markup

Compare and explore

faq

A100 40GB rental questions

How much does it cost to rent an A100 40GB in the cloud?

A100 40GB cloud instances on GPU.ai currently start at $0.670 per GPU per hour, billed per second. That is the live floor price across every provider in our fleet. We route your launch to the cheapest provider with stock and add zero markup.

How much VRAM does the A100 40GB have?

The NVIDIA A100 40GB has 40 GB of HBM2 memory with up to 1,555 GB/s of bandwidth. GPU.ai reports the minimum VRAM actually delivered across our fleet, so the figure you see at launch is the figure you get.

Can I rent a multi-GPU A100 40GB node?

Yes: real 4× A100 40GB nodes are launchable right now, from $0.818 per GPU per hour. We only advertise node sizes that exist in live inventory, not theoretical configurations.

How does GPU.ai's A100 40GB pricing compare to AWS, Azure, and CoreWeave?

GPU.ai aggregates a dozen-plus GPU clouds and passes through each provider's raw price with no markup, so the same silicon is typically far cheaper than hyperscaler list rates. The comparison table on this page shows the current public on-demand list prices side by side.

How does billing work?

Billing is per second, starting when your instance is SSH-ready and stopping the moment you terminate. There are no minimum commitments, no reservation fees, and no egress surprises: you pay the listed hourly rate pro-rated to the second.