Fal logo

GPU Cloud Provider

Fal

GPUs
6
Uptime SLA
99.99% uptime

GPU Marketplace

Estimates onlyrates are collected automatically from public provider pages and may be out of date. Prices vary by region, commitment term, and availability, and typically exclude storage, egress, and tax. Confirm current pricing with the provider before purchasing. Last collected .

Company Profile

Company TypeGenerativemedia platformfor developers.
Provider TypeGenerativemedia platformfor developers.
Legal Entityfal

Infrastructure

GPU FleetAccess 1000s of H100, H200 and B200 VMs with fal compute.
Total GPU CapacityAccess 1000s of H100, H200 and B200 VMs with fal compute.
developersenterprisefrontier research labs

Compute & Deployment

On-DemandOn-demand, serverless GPUs.
Reserved InstancesUsage-based or reserved pricing
VM-BasedAccess 1000s of H100, H200 and B200 VMs with fal compute.
Serverless GPUfal's globally distributed serverless engine

GPU Hardware

Latest GenH100, H200, B200
Pool Size1000s of H100, H200 and B200 VMs

Pricing Model

SubscriptionPay-Per-Use
Public Pricinghttps://fal.ai/pricing
Hidden FeesScale without lock-in or hidden fees.
Pay-as-you-goPay only for what you use. Choose per-output pricing for Serverless, or hourly GPU pricing with Compute.

Performance & Scaling

Multi-Node TrainingRun large scale training workloads
Max Cluster SizeScale from zero to thousands of GPUs instantly.
Elastic ScalingScale from zero to thousands of GPUs instantly.
Auto ScalingNo GPUs to configure, no cold starts, no autoscaler setup.
SLA99.99% uptime
Perf IsolationSpin up dedicated compute to fine-tune, train, or run custom models with guaranteed performance.

Developer Experience

Model MarketplaceChoose from 1,000+ production ready image, video, audio and 3D models.
DocumentationAPI and SDKs

Security & Compliance

SOC 2 complianceSingle Sign-OnPrivate endpointsUsage analytics24/7 priority supportNamed customers (CanvaPerplexityQuora).

Data Center Locations

Coverage

global regions

Compliance Regions

Datacenter Locations

Key Strengths

Fast inference engine (up to 10x faster)
serverless on-demand GPUs with no cold starts
large model gallery (1000+ generative media models)
dedicated clusters for training with guaranteed performance and access to modern NVIDIA hardware
enterprise features including SOC 2 compliance
Single Sign-On, private endpoints, usage analytics, and 24/7 priority support.

Known Limitations

No explicit spot instance pricing mentioned
No free tier or free credits mentioned
No SDK language list provided
No Jupyter notebook support mentioned
No data center tier or PUE rating stated.

Additional Information

Core Proposition

Fastest inference engine for diffusion models

Notable Customers

Canva
Perplexity
Quora
Poe
Trusted by over 1
500
000 developers and leading companies.
Last updated August 2026. Information subject to change.