NVIDIA

RTX A6000

RTX A6000 — illustration of the card's form factor
VRAM
48GB
FP32 TFLOPS
38.7 TFLOPS
CUDA Cores
10,752
TDP
300 W

Provider Marketplace

Cheapest
$0.30/hour
Starting from
Best Value
$0.45/hour
Starting from
Enterprise Choice
$1.67/hour
Starting from

All Cloud Providers

10 Options available
Verda logo
VerdaCheapest
Spot · preemptible
$0.30/ hour
Estimated Cost
Provision
RunPod logo
On-Demand
$0.33/ hour
Estimated Cost
Provision
$0.35/ hour
Estimated Cost
Provision
Hyperstack logo
Reserved · commitment
$0.35/ hour
Estimated Cost
Provision
$0.41/ hour
Estimated Cost
Provision
TensorDock logo
On-Demand
$0.45/ hour
Estimated Cost
Provision
Oblivus logo
On-Demand
$0.55/ hour
Estimated Cost
Provision
$0.55/ hour
Estimated Cost
Provision
Runcrate logo
On-Demand
$0.63/ hour
Estimated Cost
View Provider
Paperspace logo
Reserved · commitment
$1.67/ hour
Estimated Cost
Provision

Estimates onlyrates are collected automatically from public provider pages and may be out of date. Prices vary by region, commitment term, and availability, and typically exclude storage, egress, and tax. Confirm current pricing with the provider before purchasing. Last collected .

Compute Performance

FP3238.7 TFLOPS

Architecture

MicroarchitectureAmpere
CUDA Cores10752
Tensor CoresThird-generation, 336 Tensor Cores
RT CoresSecond-generation, 84 RT Cores
Sparse AccelerationSupported (structural sparsity)
Dynamic PrecisionSupported (TF32)

Memory & VRAM

Memory TypeGDDR6
Total Capacity48GB
Bandwidth768 GB/s
Bus Width384-bit
ECC SupportYes
Memory PoolingNVLink (share GPU performance and memory)

Connectivity & Scaling

InterconnectNVLink
GenerationThird-generation NVLink
IB Bandwidth112.5 GB/s
PCIe InterfacePCIe 4.0 xx16
Topology2-way NVLink
P2P MemoryYes (single scalable memory via NVLink)

Virtualization

vGPU ReadinessSupported (NVIDIA vPC/vApps, NVIDIA RTX Virtual Workstation, NVIDIA Virtual Compute Server)
GPU SharingvGPU software, NVLink
Virt EfficiencyNear bare-metal (vendor claim)

Power & Efficiency

TDP300 W
Peak Power300
Connectors1x 8-pin CPU
Thermal LimitsActive

Physical Design

Form FactorPCIe
FHFLYes
Slot WidthDual-slot
CoolingActive

Software Ecosystem

CUDACUDA 11.6
PyTorchTested with PyTorch (benchmark)
Driver StabilityNVIDIA RTX Enterprise drivers continually tuned for software compatibility

Server & Deployment

OEM AvailabilityFeaturing a dual-slot, power efficient design, the RTX A6000 is up to 2X more power efficient than Turing GPUs and crafted to fit into a wide range of workstations from worldwide OEM vendors.

System Compatibility

Required PCIePCI Express Gen 4 x16
Rack PowerTotal board power: 300 W
OS CompatWindows 10, Windows 11, and Linux

Benchmarks & Throughput

Structured Sparsity

Hardware support for structural sparsity doubles the throughput for inferencing.

Training Benchmarks

New Tensor Float 32 (TF32) precision provides up to 5X the training throughput over the previous generation to accelerate AI and data science model training without requiring any code changes.

Scaling Efficiency

With up to 112 gigabytes per second (GB/s) of bidirectional bandwidth and combined graphics memory of up to 96GB, professionals can tackle the largest rendering, AI, virtual reality, and visual computing workloads.

Multi-GPU Scalability

Scaling Efficiency

Single GPUup to 2X more power efficient than Turing GPUs
2-GPUNVLink bandwidth: 112.5 GB/s (bidirectional); combined GPU memory (2 GPUs): up to 96 GB

Scaling Characteristics

ParallelismNVLink (multi-GPU scaling, connects two RTX A6000s); NVIDIA virtual GPU (vGPU) software support; Four-way explicit VR SLI; NVIDIA GPUDirect for Video; PCI Express Gen 4

Workload Readiness

LLM Training

up to 5X the training throughput over the previous generation

LLM Inference

doubles the throughput for inferencing

HPC / Simulation

significant performance improvements for graphics and simulation workflows (including CAE)

Scientific Computing

designed for scientists to meet compute-intensive workflows

Market Authority

Community Benchmarks

SPECviewperf 2020; Autodesk VRED; BERT Large Training

Enterprise Cases

Predator Cycling; Archilime; David Baylis (Real-Time Automotive Rendering)

Key Strengths

Limitations

Expert Insight

The RTX A6000 represents a strategic leap in AI compute. When comparing cloud providers, consider not just the hourly rate, but also the interconnect bandwidth (InfiniBand/NVLink) and regional availability which can significantly impact total cost of ownership for large-scale training.

Glossary Terms

FP32 TFLOPS
VRAM
TDP
Cores
Information updated daily. Cloud pricing subject to vendor availability.