NVIDIA

Quadro RTX 5000

Quadro RTX 5000 — illustration of the card's form factor
VRAM
16GB
FP32 TFLOPS
11.2 TFLOPS
CUDA Cores
3,072
TDP
265 W

Provider Marketplace

Cheapest
$0.07/hour
Starting from
Best Value
Awaiting listings
No additional provider yet
Enterprise Choice
$0.66/hour
Starting from

All Cloud Providers

2 Options available
OVHcloud logo
OVHcloudCheapest
On-Demand
$0.07/ hour
Estimated Cost
Provision
Paperspace logo
Reserved · commitment
$0.66/ hour
Estimated Cost
Provision

Estimates onlyrates are collected automatically from public provider pages and may be out of date. Prices vary by region, commitment term, and availability, and typically exclude storage, egress, and tax. Confirm current pricing with the provider before purchasing. Last collected .

Compute Performance

FP3211.2 TFLOPS

Architecture

MicroarchitectureTuring
CUDA Cores3072
Tensor Cores384 Tensor Cores
RT Cores48 RT Cores

Memory & VRAM

Memory TypeGDDR6
Total Capacity16GB
Bandwidth448 GB/s
Bus Width256-bit
ECC SupportYes
Memory PoolingNVLink memory pooling

Connectivity & Scaling

InterconnectNVLink
IB Bandwidth50 GB/s (bidirectional)
PCIe InterfacePCI Express 3.0 xx16
TopologyNVLink (connects two GPUs)
GPUDirect RDMAYes

Virtualization

GPU SharingNVIDIA GPUDirect; NVLink (connects 2 GPUs)

Power & Efficiency

TDP265 W
Thermal LimitsActive

Physical Design

Form FactorPCIe
Slot WidthDual-slot
CoolingActive

Software Ecosystem

Compiler StackCUDA, DirectCompute, OpenCL ™
Driver StabilityQuadro cards are certified with a broad range of sophisticated professional applications, tested by leading workstation manufacturers, and backed by a global team of support specialists.

Server & Deployment

OEM AvailabilityBUY FROM PARTNERS
PreconfiguredLenovo ThinkPad and ThinkStation P Series; HP Z Workstation Family; Boston Limited server, storage and workstation solutions

System Compatibility

Required PCIePCI Express 3.0 x 16
MotherboardRequires a PCI Express 3.0 x16 slot; full-height, dual-slot card (4.4” H x 10.5” L)
Rack PowerTotal board power: 265 W; Total graphics power: 230 W
OS CompatWindows 7, 8, 8.1, 10 and Linux

Benchmarks & Throughput

Scaling Efficiency

Connect two Quadro RTX 5000s together with up to 50 GB/s of bandwidth for a combined 32 GB of GDDR6 memory to tackle larger rendering, AI, virtual reality, or visualisation workloads.

Multi-GPU Scalability

Scaling Efficiency

Single GPU11.2 TFLOPS
2-GPU50 GB/s (bidirectional)

Scaling Characteristics

ParallelismCUDA, DirectCompute, OpenCL

Workload Readiness

Market Authority

Key Strengths

Limitations

Expert Insight

The Quadro RTX 5000 represents a strategic leap in AI compute. When comparing cloud providers, consider not just the hourly rate, but also the interconnect bandwidth (InfiniBand/NVLink) and regional availability which can significantly impact total cost of ownership for large-scale training.

Glossary Terms

FP32 TFLOPS
VRAM
TDP
Cores
Information updated daily. Cloud pricing subject to vendor availability.