GPU.ai has a new look. The compute is the same.Read about the redesign  ↗

gpu catalog: t4

Rent T4●

The classic low-power inference card: 70 watts and 16 GB, still everywhere and still fine for embeddings and classic ML serving.

Live platform rates

● 29 GPU types online
GPUVRAMRegionsAvailabilityPrice / GPU
T416G—out of stock—

T4 is out of stock right now — these are live:

The classic low-power inference card: 70 watts and 16 GB, still everywhere and still fine for embeddings and classic ML serving.

VRAM16 GB GDDR6
Memory bandwidth0.3 TB/s
FP16 tensor65 TFLOPS
ArchitectureTuring
InterconnectPCIe 3.0
Board power70 W
CUDA cores2,560
Released2018
RTX 5090community32G8 regionshigh availability$0.47/hrlaunch
L40S48G3 regionshigh availability$0.80/hrlaunch
A10080G5 regionshigh availability$1.07/hrlaunch
RTX 4090community24G7 regionshigh availability$0.34/hrlaunch
RTX PRO 6000community96G3 regionshigh availability$1.69/hrlaunch
RTX PRO 600096G3 regionshigh availability$2.09/hrlaunch
H200 SXMcommunity141G4 regionshigh availability$3.59/hrlaunch
RTX 3090community24G6 regionsavailable$0.15/hrlaunch
A100community80G6 regionsavailable$0.95/hrlaunch
H100 SXM80G4 regionsavailable$3.49/hrlaunch
H200 SXM141G4 regionsavailable$4.59/hrlaunch
A100 40GBcommunity40G5 regionsavailable$0.47/hrlaunch
RTX 6000 ADA48G2 regionsavailable$0.65/hrlaunch
L40Scommunity48G3 regionsavailable$0.79/hrlaunch
RTX 3080community10G4 regionsavailable$0.11/hrlaunch
RTX A400016G1 regionavailable$0.12/hrlaunch
RTX A600048G3 regionsavailable$0.40/hrlaunch
RTX 5080community16G6 regionsavailable$0.23/hrlaunch
RTX A4000community16G4 regionsavailable$0.10/hrlaunch
L424G3 regionsavailable$0.49/hrlaunch
RTX 4080community16G4 regionsrunning low$0.21/hrlaunch
H100 PCIE80G2 regionsrunning low$1.98/hrlaunch
RTX 2000 ADA16G2 regionsrunning low$0.24/hrlaunch
A4048G2 regionsrunning low$0.49/hrlaunch
B300288G2 regionsrunning low$7.89/hrlaunch
L4048G1 regionrunning low$0.78/hrlaunch
RTX A6000community48G2 regionsrunning low$0.40/hrlaunch
H100 NVLcommunity94G4 regionsrunning low$2.67/hrlaunch
RTX 4000 ADAcommunity20G1 regionrunning low$0.20/hrlaunch
RTX A500024G1 regionrunning low$0.27/hrlaunch
RTX 4000 ADA20G1 regionrunning low$0.28/hrlaunch
L4community24G3 regionsrunning low$0.33/hrlaunch
RTX 309024G1 regionrunning low$0.50/hrlaunch
L40community48G1 regionrunning low$0.69/hrlaunch
RTX PRO 450032G1 regionrunning low$0.72/hrlaunch
RTX 6000 ADAcommunity48G1 regionrunning low$0.74/hrlaunch
RTX 509032G1 regionrunning low$0.99/hrlaunch
H100 NVL94G1 regionrunning low$3.19/hrlaunch
H200 NVLcommunity141G2 regionsrunning low$5.47/hrlaunch
B200community192G2 regionsrunning low$7.50/hrlaunch
H100 SXMcommunity80G2 regionsrunning low$3.56/hrlaunch
V100community16G1 regionrunning low$0.13/hrlaunch
RTX A4500community20G1 regionrunning low$0.13/hrlaunch
RTX A5000community24G1 regionrunning low$0.30/hrlaunch
H200 NVL141G1 regionrunning low$3.29/hrlaunch

lowest live on-demand rate per GPU, refreshed continuously · per-second billing

T4 price history

last 90 days

Our floor price for the T4, sampled continuously across every provider in the fleet and rolled up to one point per day. A dot on the zero line is a day we checked and found nothing launchable; a shaded break is a day we did not sample. Those are different facts, so this chart never draws one as the other.

We have not collected enough daily samples for the T4 yet. The trend appears here once history accumulates — we do not draw a line through data we never had.

Compare and explore

faq

T4 rental questions●

How much does it cost to rent a T4 in the cloud?

T4 capacity is temporarily out of stock on GPU.ai. When available, pricing is the provider's raw rate with zero markup, billed per second. Check this page again. Stock and prices refresh continuously.

How much VRAM does the T4 have?

The NVIDIA Tesla T4 has 16 GB of GDDR6 memory with up to 320 GB/s of bandwidth. GPU.ai reports the minimum VRAM actually delivered across our fleet, so the figure you see at launch is the figure you get.

How does GPU.ai's T4 pricing compare to AWS, Azure, and CoreWeave?

GPU.ai aggregates a dozen-plus GPU clouds and passes through each provider's raw price with no markup, so the same silicon is typically far cheaper than hyperscaler list rates. The comparison table on this page shows the current public on-demand list prices side by side.

How does billing work?

Billing is per second, starting when your instance is SSH-ready and stopping the moment you terminate. There are no minimum commitments, no reservation fees, and no egress surprises: you pay the listed hourly rate pro-rated to the second.