RunPod logo

GPU Cloud Provider · Unknown

RunPod

RunPod offers on-demand GPU clusters optimized for AI, ML, large language models (LLMs), and high-performance computing (HPC) workloads. They provide instant deployment, on-demand scalability, and flexible billing measured by the second, without minimum commitments or contracts.

GPUs
47
Founded
Unknown
Countries
4

GPU Marketplace

$0.16/hour
$0.27/hour
$0.33/hour
NVIDIA A40On-Demand
$0.35/hour
NVIDIA L4On-Demand
$0.44/hour
NVIDIA A40On-Demand
$0.44/hour
NVIDIA L4On-Demand
$0.49/hour
$0.53/hour
NVIDIA L40On-Demand
$0.69/hour
NVIDIA L40SOn-Demand
$0.79/hour
NVIDIA L40On-Demand
$0.82/hour
NVIDIA L40SOn-Demand
$0.99/hour
$1.99/hour
$2.59/hour
$2.69/hour
$2.89/hour
$3.19/hour
$3.29/hour
$3.59/hour
$4.31/hour
$4.59/hour
$5.93/hour
$5.98/hour
$6.79/hour
$6.94/hour
$7.89/hour
$8.64/hour
$9.98/hour

Estimates onlyrates are collected automatically from public provider pages and may be out of date. Prices vary by region, commitment term, and availability, and typically exclude storage, egress, and tax. Confirm current pricing with the provider before purchasing. Last collected .

Company Profile

Company TypeOrganization
Provider TypeAI Developer Cloud
Legal EntityRunpod Inc.

Infrastructure

GPU Fleet30+ GPU models, examples include A100, H100, RTX 6000 Ada, L4/L40 series
Total GPU CapacityThousands of GPUs across 30+ regions
Network FabricInfiniBand, RoCE v2
Connectivity1,600–3,200 Gbps
StorageNetwork Storage with shared filesystems
AvailabilityGA
developers, researchers, and AI companies

Compute & Deployment

On-DemandRent on-demand cloud GPU instances
Spot / InterruptibleWe offer spot instances where GPU capacity is available at a discount, but with the risk of eviction when demand spikes.
Reserved InstancesPrefer renting GPU capacity on a reserved basis? Get discounted rates on active and flex instances.
Bare MetalDedicated GPU clusters with guaranteed availability, custom configurations, SLA-backed uptime, and discounted rates for enterprises scaling to 10,000+ GPUs.
VM-BasedGPU Pods are dedicated GPU instances you can spin up on Runpod. Unlike abstracted serverless GPUs, Pods give you full control over the underlying VM, drivers, and environment.
Container-BasedGPU Pods support custom Docker images.
Serverless GPUServerless for API inference
Spin-Up TimeSpin up any GPU instance in under 30 seconds.

GPU Hardware

Latest GenH100, A100, RTX 4090
Multi-GPU NodesPersistent storage, multi-GPU support, no babysitting required.
Pool SizeThousands of GPUs across 30+ regions.
NVLinkH100 NVL
PCIe vs SXMH100 PCIe; H100 SXM; A100 PCIe; A100 SXM

Pricing Model

Per HourPricing is shown as an hourly rate but billed by the millisecond. You only pay for the exact time your Pod runs.
Per MinuteGPU rental billed by the second. No egress fees, no minimums, no surprises. Run a job for 3 minutes. Pay for 3 minutes.
SubscriptionOn-demand GPU cloud pricing with no long-term commitments. Rent by the second or lock in savings with reservations.
Reserved DiscountPrefer renting GPU capacity on a reserved basis? Get discounted rates on active and flex instances. No per-second premium, just predictable cloud GPU pricing.
Public Pricinghttps://www.runpod.io/pricing
Hidden FeesGPU rental billed by the second. No egress fees, no minimums, no surprises.
Egress ChargesNo fees for ingress/egress. Persistent and temporary storage available.
Pay-as-you-goOn-demand GPU cloud pricing with no long-term commitments. Rent by the second or lock in savings with reservations.

Performance & Scaling

Multi-Node TrainingA100 and H100 SXM GPU rentals built for long training runs. Persistent storage, multi-GPU support, no babysitting required.
Max Cluster SizeLaunch multi-GPU clusters in minutes with no commitments. Scale up to 64 GPUs, attach shared storage, and pay only for what you use.
Elastic ScalingOn-demand GPU instances that spin up fast and scale with your workload. No cold starts, no idle costs.
Perf IsolationCloud GPU instances are dedicated GPU environments for AI development, training, fine-tuning, batch jobs, and long-running workloads.

Developer Experience

OnboardingSelf-serve signup; No credit card required to explore; No minimums once you deploy; Sign up and deploy in the console (Pods → Deploy); instances live in under 30 seconds
FrameworksAll major AI frameworks compatible via Docker
CLI ToolingCLI & SDKs; manage from terminal and CI pipelines
TemplatesBase images with common ML stacks
Model MarketplacePublic Endpoints / Runpod Hub: instant access to pre-deployed AI models via API
DocumentationAPI reference and automation (Full API access); CLI & SDKs; GitHub & CI/CD deployment integration; FAQs with technical answers; base images with common ML stacks
API FeaturesSelf-service provisioning through intuitive console

Security & Compliance

SecuritySOC2 Type II compliance
ComplianceSOC 2 Type II
Claim: 1M+ Developers chose RunpodClient logos: Wix, Otovo, Scatter Lab, Abzu, Aneta, Perplexity, Replit, CivitailegalName: Runpod Inc.support contact email: help@runpod.iofoundingDate: 2022-10-31

Data Center Locations

Coverage

CountriesUS, Europe, Asia, Australia
31 regions across the USEuropeAsia, and Australia

Compliance Regions

Datacenter Locations

Key Strengths

Fast 30-second deploys
31 global regions
30+ GPU models including H100/A100/RTX 4090
per-second/millisecond billing
full API/CLI/SDKs and CI/CD integration
persistent storage with no ingress/egress fees
spot instances and reservations
Reserved Clusters with guaranteed availability and SLA-backed uptime
self-serve exploration with no credit card required

Additional Information

Support Options

Not specified

Community

["LinkedIn","X (Twitter)","Public GitHub org","YouTube channel","Community Cloud"]

Core Proposition

Rent on-demand cloud GPU instances with H100, A100, RTX 4090, and 30+ GPU models across 31 global regions, with per-second billing for AI workloads.

Notable Customers

Civit AI
Cognition
Cursor
Hugging Face
Magic
Otovo
Perplexity
Replit
Wix
Scatter Lab
Abzu
Aneta
Civitai
Last updated March 2026. Information subject to change.