

A40 Specifications
Ampere architecture with dedicated ray tracing cores, Tensor Cores, and 48 GB GDDR6 for professional visualization and lightweight AI workloads.
48 GB
GDDR6 Memory
Large frame buffer for complex scenes and datasets
696 GB/s
Memory Bandwidth
Sustained throughput for rendering and simulation
84
RT Cores
Second-gen hardware ray tracing at 73 TFLOPS
149.7
FP16 TFLOPS
299.4 TFLOPS with sparsity enabled
NVIDIA B200 on Denvr AI Cloud

1T Parameter Training
Scale across 8-GPU NVLink nodes for 1,440 GB of total VRAM and 1.8 TB/s per-GPU interconnect.

vGPU Workstations
Powerful virtual workstation instances for remote users, enabling high-end remote design, AI, and compute workloads.

Lightweight Inference
Serve smaller models and embedding pipelines where HBM bandwidth isn't required. A40 delivers capable inference at a lower price point than A100.

Managed Storage
High-performance Weka filesystem and local NVMe available. No external storage to provision for datasets, checkpoints, or model artifacts.
Configurations
Per-minute billing with on-demand and reserved options. All configurations available as bare metal, VM, or model endpoints.
Platform
GPUs
On-Demand
VRAM
vCPUs
Memory
Local Storage
Interconnect
NVIDIA A40
—
—
—
48 GB
—
—
—
PLATFORM
GPU COUNT
GPU VRAM
vCPUs
MEMORY
LOCAL STORAGE
INTERCONNECT
GPU Count
GPU VRAM
vCPUs
Memory
Local Storage
Interconnect
GPU Count
Related GPUs
Compare Denvr GPU options by workload and performance requirements.
Intel Gaudi 2
VRAM
VRAM
VRAM
VRAM
Memory Bandwidth
VRAM
VRAM
VRAM
FP64/FP32
VRAM
VRAM
VRAM
FP16
VRAM
VRAM
VRAM
FP8
VRAM
VRAM
VRAM
NVLink
VRAM
VRAM
VRAM
On-Demand Pricing
VRAM
VRAM
VRAM
PLATFORM
VRAM
BANDWIDTH
FP64/FP32
FP16
FP8
NVLINK
GPU Count
GPU VRAM
vCPUs
Memory
Local Storage
Interconnect
GPU Count
PLATFORM
VRAM
BANDWIDTH
FP64/FP32
FP16
FP8
NVLINK
GPU Count
GPU VRAM
vCPUs
Memory
Local Storage
Interconnect
GPU Count
PLATFORM
VRAM
BANDWIDTH
FP64/FP32
FP16
FP8
NVLINK
GPU Count
GPU VRAM
vCPUs
Memory
Local Storage
Interconnect
GPU Count
Infrastructure you can trust at scale
As an NVIDIA Cloud Partner we build and operate AI clusters following NVIDIA Reference Architectures. Your models and data are supported via strict privacy safeguards and SOC 2 Type 2 security practices.









