NVIDIA
P4

Provider Marketplace
All Cloud Providers
Estimates only — rates are collected automatically from public provider pages and may be out of date. Prices vary by region, commitment term, and availability, and typically exclude storage, egress, and tax. Confirm current pricing with the provider before purchasing. Last collected .
Compute Performance
Architecture
Memory & VRAM
Connectivity & Scaling
Power & Efficiency
Physical Design
Thermals & Cooling
Software Ecosystem
System Compatibility
Benchmarks & Throughput
Inference Benchmarks
22 TOPs of INT8 inference; slashes latency by 15X
Multi-GPU Scalability
Scaling Efficiency
Scaling Characteristics
Workload Readiness
LLM Inference
22 TOPs (INT8)
Scientific Computing
Single-Precision Performance 5.5 TeraFLOPS
Real-Time Serving
Hardware-decode engine capable of transcoding and inferencing 35 HD video streams in real time
Market Authority
Key Strengths
Limitations
Also in the Lineup
Expert Insight
The P4 represents a strategic leap in AI compute. When comparing cloud providers, consider not just the hourly rate, but also the interconnect bandwidth (InfiniBand/NVLink) and regional availability which can significantly impact total cost of ownership for large-scale training.