NVIDIA

T4

T4 — illustration of the card's form factor
VRAM
16GB
FP32 TFLOPS
8.1 TFLOPS
CUDA Cores
2,560
TDP
70 W

Provider Marketplace

Cheapest
$0.13/hour
Starting from
Best Value
$0.42/hour
Starting from
Enterprise Choice
$2.38/hour
Starting from

All Cloud Providers

9 Options available
RedSwitches logo
RedSwitchesCheapest
On-DemandSingaporeSingapore
$0.13/ hour
Estimated Cost
Provision
Google Cloud logo
Reserved · commitment
$0.16/ hour
Estimated Cost
Provision
Tencent Cloud logo
On-Demand
$0.20/ hour
Estimated Cost
Provision
$0.33/ hour
Estimated Cost
Provision
Lightning AI logo
On-Demand
$0.42/ hour
Estimated Cost
Provision
Modal logo
On-Demand
$0.59/ hour
Estimated Cost
Provision
Baseten logo
On-Demand
$0.63/ hour
Estimated Cost
Provision
Replicate logo
On-Demand
$0.81/ hour
Estimated Cost
Provision
Alibaba Cloud logo
On-DemandChina (Beijing),China (Shanghai),Singapore
$2.38/ hour
Estimated Cost
Provision

Estimates onlyrates are collected automatically from public provider pages and may be out of date. Prices vary by region, commitment term, and availability, and typically exclude storage, egress, and tax. Confirm current pricing with the provider before purchasing. Last collected .

Compute Performance

FP328.1 TFLOPS
FP1665 TFLOPS
INT8130 TOPS
INT4260 TOPS

Architecture

MicroarchitectureTuring
CUDA Cores2560
Tensor CoresTuring Tensor Cores, 320
Dynamic PrecisionSupported (FP32/FP16/INT8/INT4)

Memory & VRAM

Memory TypeGDDR6
Total Capacity16GB
Bandwidth320+ GB/s
ECC SupportYes

Connectivity & Scaling

InterconnectPCIe
GenerationPCIe Gen3
IB Bandwidth32 GB/sec
PCIe InterfaceGen3 xx16

Virtualization

vGPU ReadinessSupported (NVIDIA Virtual Compute Server (vCS))
GPU SharingVirtualization via NVIDIA Virtual Compute Server (vCS)

Power & Efficiency

TDP70 W
Thermal LimitsThermal Solution Passive
Efficiencyproviding an incredible 50X higher energy efficiency compared to CPUs

Physical Design

Form FactorLow-Profile PCIe
CoolingPassive
Rack DensityOptimized for scale-out servers

Thermals & Cooling

DC Heat70 watts

Server & Deployment

Rack-ScaleOptimized for scale-out servers

System Compatibility

Required PCIex16 PCIe Gen3
MotherboardLow-Profile PCIe (small PCIe form factor)

Multi-GPU Scalability

Scaling Efficiency

Single GPU50X higher energy efficiency compared to CPUs

Scaling Characteristics

ParallelismCUDA, NVIDIA TensorRT, ONNX; GPU virtualization via NVIDIA Virtual Compute Server (vCS)

Workload Readiness

Real-Time Serving

up to 40X times better throughput

Market Authority

Key Strengths

Limitations

Expert Insight

The T4 represents a strategic leap in AI compute. When comparing cloud providers, consider not just the hourly rate, but also the interconnect bandwidth (InfiniBand/NVLink) and regional availability which can significantly impact total cost of ownership for large-scale training.

Glossary Terms

FP32 TFLOPS
VRAM
TDP
Cores
Information updated daily. Cloud pricing subject to vendor availability.