GPU Marketplace
Estimates only — rates are collected automatically from public provider pages and may be out of date. Prices vary by region, commitment term, and availability, and typically exclude storage, egress, and tax. Confirm current pricing with the provider before purchasing. Last collected .
Company Profile
Company TypeGenerativemedia platformfor developers.
Provider TypeGenerativemedia platformfor developers.
Legal Entityfal
Infrastructure
GPU FleetAccess 1000s of H100, H200 and B200 VMs with fal compute.
Total GPU CapacityAccess 1000s of H100, H200 and B200 VMs with fal compute.
developersenterprisefrontier research labs
Compute & Deployment
On-DemandOn-demand, serverless GPUs.
Reserved InstancesUsage-based or reserved pricing
VM-BasedAccess 1000s of H100, H200 and B200 VMs with fal compute.
Serverless GPUfal's globally distributed serverless engine
GPU Hardware
Latest GenH100, H200, B200
Pool Size1000s of H100, H200 and B200 VMs
Pricing Model
SubscriptionPay-Per-Use
Public Pricinghttps://fal.ai/pricing
Hidden FeesScale without lock-in or hidden fees.
Pay-as-you-goPay only for what you use. Choose per-output pricing for Serverless, or hourly GPU pricing with Compute.
Performance & Scaling
Multi-Node TrainingRun large scale training workloads
Max Cluster SizeScale from zero to thousands of GPUs instantly.
Elastic ScalingScale from zero to thousands of GPUs instantly.
Auto ScalingNo GPUs to configure, no cold starts, no autoscaler setup.
SLA99.99% uptime
Perf IsolationSpin up dedicated compute to fine-tune, train, or run custom models with guaranteed performance.
Developer Experience
Model MarketplaceChoose from 1,000+ production ready image, video, audio and 3D models.
DocumentationAPI and SDKs
Security & Compliance
SOC 2 complianceSingle Sign-OnPrivate endpointsUsage analytics24/7 priority supportNamed customers (CanvaPerplexityQuora).
Data Center Locations
Coverage
global regions
Compliance Regions
Datacenter Locations
Key Strengths
Fast inference engine (up to 10x faster)
serverless on-demand GPUs with no cold starts
large model gallery (1000+ generative media models)
dedicated clusters for training with guaranteed performance and access to modern NVIDIA hardware
enterprise features including SOC 2 compliance
Single Sign-On, private endpoints, usage analytics, and 24/7 priority support.
Known Limitations
No explicit spot instance pricing mentioned
No free tier or free credits mentioned
No SDK language list provided
No Jupyter notebook support mentioned
No data center tier or PUE rating stated.
Additional Information
Core Proposition
Fastest inference engine for diffusion models
Notable Customers
Canva
Perplexity
Quora
Poe
Trusted by over 1
500
000 developers and leading companies.
Last updated August 2026. Information subject to change.


