We match your deposits, up to $150 in free credits. Ends October 1.Learn more  

gpu catalog: t4

Rent T4

The classic low-power inference card: 70 watts and 16 GB, still everywhere and still fine for embeddings and classic ML serving.

Live platform rates

27 GPU types online
GPUVRAMRegionsAvailabilityPrice / GPU
T416Gout of stock

T4 is out of stock right now — these are live:

The classic low-power inference card: 70 watts and 16 GB, still everywhere and still fine for embeddings and classic ML serving.

VRAM16 GB GDDR6
Memory bandwidth0.3 TB/s
FP16 tensor65 TFLOPS
ArchitectureTuring
InterconnectPCIe 3.0
Board power70 W
CUDA cores2,560
Released2018
L40S48G4 regionshigh availability$0.80/hrlaunch
RTX PRO 600096G5 regionshigh availability$1.99/hrlaunch
RTX 6000 ADA48G2 regionshigh availability$0.65/hrlaunch
RTX 5090community32G8 regionshigh availability$0.35/hrlaunch
L4048G2 regionshigh availability$0.78/hrlaunch
RTX 4090community24G8 regionshigh availability$0.27/hrlaunch
RTX A600048G2 regionshigh availability$0.45/hrlaunch
A10080G7 regionshigh availability$1.07/hrlaunch
RTX PRO 450032G2 regionshigh availability$0.72/hrlaunch
H100 SXM80G5 regionshigh availability$3.49/hrlaunch
A100community80G5 regionshigh availability$0.88/hrlaunch
H200 SXM141G6 regionshigh availability$3.59/hrlaunch
H200 SXMcommunity141G5 regionshigh availability$3.59/hrlaunch
RTX 3090community24G6 regionshigh availability$0.12/hrlaunch
L40Scommunity48G4 regionsavailable$0.79/hrlaunch
H100 PCIE80G2 regionsavailable$1.98/hrlaunch
RTX A4000community16G5 regionsavailable$0.08/hrlaunch
A100 40GBcommunity40G6 regionsavailable$0.56/hrlaunch
RTX 5080community16G7 regionsavailable$0.16/hrlaunch
RTX 409024G3 regionsavailable$0.74/hrlaunch
RTX 509032G2 regionsavailable$0.99/hrlaunch
L424G3 regionsavailable$0.44/hrlaunch
RTX 3080community10G5 regionsavailable$0.10/hrlaunch
H200 NVLcommunity141G3 regionsavailable$3.98/hrlaunch
RTX 6000 ADAcommunity48G2 regionsavailable$0.74/hrlaunch
RTX PRO 6000community96G2 regionsavailable$1.69/hrlaunch
B300288G2 regionsavailable$7.89/hrlaunch
H100 PCIEcommunity80G2 regionsavailable$2.54/hrlaunch
RTX 2000 ADA16G2 regionsrunning low$0.24/hrlaunch
A4048G2 regionsrunning low$0.49/hrlaunch
H100 NVL94G2 regionsrunning low$3.19/hrlaunch
B200192G2 regionsrunning low$6.79/hrlaunch
RTX A5000community24G2 regionsrunning low$0.23/hrlaunch
L4community24G2 regionsrunning low$0.27/hrlaunch
H100 SXMcommunity80G3 regionsrunning low$2.27/hrlaunch
B200community192G2 regionsrunning low$5.32/hrlaunch
RTX A6000community48G2 regionsrunning low$0.41/hrlaunch
RTX 4080community16G3 regionsrunning low$0.18/hrlaunch
H200 NVL143G1 regionrunning low$3.79/hrlaunch
A100 40GB40G1 regionrunning low$0.89/hrlaunch
V100community16G2 regionsrunning low$0.09/hrlaunch
H100 NVLcommunity94G2 regionsrunning low$2.37/hrlaunch

lowest live on-demand rate per GPU, refreshed continuously · per-second billing

Compare and explore

faq

T4 rental questions

How much does it cost to rent a T4 in the cloud?

T4 capacity is temporarily out of stock on GPU.ai. When available, pricing is the provider's raw rate with zero markup, billed per second. Check this page again. Stock and prices refresh continuously.

How much VRAM does the T4 have?

The NVIDIA Tesla T4 has 16 GB of GDDR6 memory with up to 320 GB/s of bandwidth. GPU.ai reports the minimum VRAM actually delivered across our fleet, so the figure you see at launch is the figure you get.

How does GPU.ai's T4 pricing compare to AWS, Azure, and CoreWeave?

GPU.ai aggregates a dozen-plus GPU clouds and passes through each provider's raw price with no markup, so the same silicon is typically far cheaper than hyperscaler list rates. The comparison table on this page shows the current public on-demand list prices side by side.

How does billing work?

Billing is per second, starting when your instance is SSH-ready and stopping the moment you terminate. There are no minimum commitments, no reservation fees, and no egress surprises: you pay the listed hourly rate pro-rated to the second.