GPU.ai has a new look. The compute is the same.Read about the redesign  ↗

gpu catalog: a30

Rent A30●

A data center inference card with real HBM2 bandwidth at low power, an efficient host for mid-size model serving.

Live platform rates

● 29 GPU types online
GPUVRAMRegionsAvailabilityPrice / GPU
A3024G—out of stock—

A30 is out of stock right now — these are live:

A data center inference card with real HBM2 bandwidth at low power, an efficient host for mid-size model serving.

VRAM24 GB HBM2
Memory bandwidth0.9 TB/s
FP16 tensor165 TFLOPS
ArchitectureAmpere
InterconnectPCIe 4.0
Board power165 W
Released2021
RTX 5090community32G8 regionshigh availability$0.47/hrlaunch
L40S48G3 regionshigh availability$0.80/hrlaunch
A10080G5 regionshigh availability$1.07/hrlaunch
RTX 4090community24G7 regionshigh availability$0.34/hrlaunch
RTX PRO 6000community96G3 regionshigh availability$1.69/hrlaunch
RTX PRO 600096G3 regionshigh availability$2.09/hrlaunch
H200 SXMcommunity141G4 regionshigh availability$3.59/hrlaunch
RTX 3090community24G6 regionsavailable$0.15/hrlaunch
A100community80G6 regionsavailable$0.95/hrlaunch
H100 SXM80G4 regionsavailable$3.49/hrlaunch
H200 SXM141G4 regionsavailable$4.59/hrlaunch
A100 40GBcommunity40G5 regionsavailable$0.47/hrlaunch
RTX 6000 ADA48G2 regionsavailable$0.65/hrlaunch
L40Scommunity48G3 regionsavailable$0.79/hrlaunch
RTX 3080community10G4 regionsavailable$0.11/hrlaunch
RTX A400016G1 regionavailable$0.12/hrlaunch
RTX A600048G3 regionsavailable$0.40/hrlaunch
RTX 5080community16G6 regionsavailable$0.23/hrlaunch
RTX A4000community16G4 regionsavailable$0.10/hrlaunch
L424G3 regionsavailable$0.49/hrlaunch
RTX 4080community16G4 regionsrunning low$0.21/hrlaunch
H100 PCIE80G2 regionsrunning low$1.98/hrlaunch
RTX 2000 ADA16G2 regionsrunning low$0.24/hrlaunch
A4048G2 regionsrunning low$0.49/hrlaunch
B300288G2 regionsrunning low$7.89/hrlaunch
L4048G1 regionrunning low$0.78/hrlaunch
RTX A6000community48G2 regionsrunning low$0.40/hrlaunch
H100 NVLcommunity94G4 regionsrunning low$2.67/hrlaunch
RTX 4000 ADAcommunity20G1 regionrunning low$0.20/hrlaunch
RTX A500024G1 regionrunning low$0.27/hrlaunch
RTX 4000 ADA20G1 regionrunning low$0.28/hrlaunch
L4community24G3 regionsrunning low$0.33/hrlaunch
RTX 309024G1 regionrunning low$0.50/hrlaunch
L40community48G1 regionrunning low$0.69/hrlaunch
RTX PRO 450032G1 regionrunning low$0.72/hrlaunch
RTX 6000 ADAcommunity48G1 regionrunning low$0.74/hrlaunch
RTX 509032G1 regionrunning low$0.99/hrlaunch
H100 NVL94G1 regionrunning low$3.19/hrlaunch
H200 NVLcommunity141G2 regionsrunning low$5.47/hrlaunch
B200community192G2 regionsrunning low$7.50/hrlaunch
H100 SXMcommunity80G2 regionsrunning low$3.56/hrlaunch
V100community16G1 regionrunning low$0.13/hrlaunch
RTX A4500community20G1 regionrunning low$0.13/hrlaunch
RTX A5000community24G1 regionrunning low$0.30/hrlaunch
H200 NVL141G1 regionrunning low$3.29/hrlaunch

lowest live on-demand rate per GPU, refreshed continuously · per-second billing

A30 price history

last 90 days

Our floor price for the A30, sampled continuously across every provider in the fleet and rolled up to one point per day. A dot on the zero line is a day we checked and found nothing launchable; a shaded break is a day we did not sample. Those are different facts, so this chart never draws one as the other.

  • Price ($/GPU-hr)
  • Sold out — 0 available
  • No data collected for this window

Daily trend, updated hourly · Full 15-minute history available via API

Compare and explore

faq

A30 rental questions●

How much does it cost to rent an A30 in the cloud?

A30 capacity is temporarily out of stock on GPU.ai. When available, pricing is the provider's raw rate with zero markup, billed per second. Check this page again. Stock and prices refresh continuously.

How much VRAM does the A30 have?

The NVIDIA A30 has 24 GB of HBM2 memory with up to 933 GB/s of bandwidth. GPU.ai reports the minimum VRAM actually delivered across our fleet, so the figure you see at launch is the figure you get.

How does GPU.ai's A30 pricing compare to AWS, Azure, and CoreWeave?

GPU.ai aggregates a dozen-plus GPU clouds and passes through each provider's raw price with no markup, so the same silicon is typically far cheaper than hyperscaler list rates. The comparison table on this page shows the current public on-demand list prices side by side.

How does billing work?

Billing is per second, starting when your instance is SSH-ready and stopping the moment you terminate. There are no minimum commitments, no reservation fees, and no egress surprises: you pay the listed hourly rate pro-rated to the second.