NVIDIA

H200

SXM

H200 SXM — illustration of the card's form factor
VRAM
141GB
FP32 TFLOPS
67 TFLOPS
TDP
700 W
Memory
HBM3e

Provider Marketplace

Cheapest
$2.00/hour
Starting from
Best Value
$3.40/hour
Starting from
Enterprise Choice
$6.00/hour
Starting from

All Cloud Providers

12 Options available
Verda logo
VerdaCheapest
Spot · preemptible
$2.00/ hour
Estimated Cost
Provision
Nebius logo
Spot · preemptible
$2.45/ hour
Estimated Cost
Provision
CoreWeave logo
Spot · preemptibleEUROPE
$2.58/ hour
Estimated Cost
Provision
Hyperstack logo
Reserved · commitment
$2.79/ hour
Estimated Cost
Provision
Oblivus logo
Reserved · commitment
$2.92/ hour
Estimated Cost
Provision
Taiga Cloud logo
On-Demand
$2.95/ hour
Estimated Cost
Provision
DigitalOcean logo
Reserved · commitment
$3.40/ hour
Estimated Cost
Provision
RunPod logo
On-Demand
$3.59/ hour
Estimated Cost
Provision
Jarvis Labs logo
On-Demand
$3.99/ hour
Estimated Cost
Provision
Together logo
Reserved · commitment
$3.99/ hour
Estimated Cost
Provision
Modal logo
On-Demand
$4.54/ hour
Estimated Cost
Provision
Crusoe Cloud logo
On-Demand
$6.00/ hour
Estimated Cost
Provision

Estimates onlyrates are collected automatically from public provider pages and may be out of date. Prices vary by region, commitment term, and availability, and typically exclude storage, egress, and tax. Confirm current pricing with the provider before purchasing. Last collected .

Compute Performance

FP6434 TFLOPS
FP3267 TFLOPS
TF32989 TFLOPS
FP161,979 TFLOPS
BF161,979 TFLOPS
FP83,958 TFLOPS
INT83,958 TFLOPS TOPS

Architecture

MicroarchitectureHopper
Sparse AccelerationSupported
Dynamic PrecisionSupported (FP8/FP16/BF16/TF32/INT8)

Memory & VRAM

Memory TypeHBM3e
Total Capacity141GB
Bandwidth4.8TB/s

Connectivity & Scaling

InterconnectNVIDIA NVLink; PCIe Gen5
IB Bandwidth900GB/s
PCIe InterfacePCIe Gen5

Virtualization

MIG SupportSupported
MIG PartitionsUp to 7 MIGs @18GB each
GPU SharingMIG (Multi-Instance GPUs) — Up to 7 MIGs @18GB each

Power & Efficiency

TDP700 W
Peak Power700

Physical Design

Form FactorSXM

Server & Deployment

OEM AvailabilityNVIDIA HGX™ H200 partner and NVIDIA-Certified Systems™ with 4 or 8 GPUs
PreconfiguredNVIDIA-Certified Systems™ with 4 or 8 GPUs
DGX/HGXThe NVIDIA HGX H200 features the NVIDIA H200 GPU
Rack-ScaleNVIDIA HGX™ H200 partner and NVIDIA-Certified Systems™ with 4 or 8 GPUs

System Compatibility

CPU Pairingdual Sapphire Rapids 8480
Required PCIePCIe Gen5

Benchmarks & Throughput

Structured Sparsity

With sparsity.

Inference Benchmarks

Llama2 70B inference 1.9X faster; GPT-3 175B inference 1.6X faster; throughput comparisons vs H100 SXM provided for Llama2 and GPT-3.

Scaling Efficiency

NVIDIA NVLink: 900GB/s

Multi-GPU Scalability

Scaling Characteristics

ParallelismUp to 7 MIGs @18GB each

Workload Readiness

LLM Training

1,979 TFLOPS

LLM Inference

3,958 TFLOPS

Vision Training

1,979 TFLOPS

HPC / Simulation

34 TFLOPS

Scientific Computing

67 TFLOPS

Market Authority

Community Benchmarks

["Llama2 70B Inference: 1.9X Faster","GPT-3 175B Inference: 1.6X Faster","High-Performance Computing: 110X Faster","H200 boosts inference speed by up to 2X compared to H100 GPUs when handling LLMs like Llama2","Throughput comparisons for H200 SXM vs H100 SXM (examples): Llama2 13B and Llama2 70B"]

Key Strengths

Limitations

Expert Insight

The H200 represents a strategic leap in AI compute. When comparing cloud providers, consider not just the hourly rate, but also the interconnect bandwidth (InfiniBand/NVLink) and regional availability which can significantly impact total cost of ownership for large-scale training.

Glossary Terms

FP32 TFLOPS
VRAM
TDP
Cores
Information updated daily. Cloud pricing subject to vendor availability.