https://www.denvrdata.com/?utm_campaign=XAds&utm_campaign_id=1&utm_medium=paid&utm_source=X
top of page
Product page - Header image.jpg

NVIDIA H100

Choose from single-GPU, to 8-GPU NVLink, or scale-out training clusters with 3,200G InfiniBand.

b200_edited.png

H100 Specifications

Fourth-generation Tensor Cores speed up all precisions, including FP64, TF32, FP32, FP16, INT8, and FP8, to reduce memory usage and increase performance while still maintaining accuracy for LLMs.

80 GB

HBM3 Memory

Run 70B+ parameter models on a single GPU.

900 GB/s

NVLink Bandwidth

Fourth-gen NVLink for multi-GPU scaling.

1,979

TFLOPS FP16

3,958 TFLOPS at FP8 with Transformer Engine.

9X

Faster Pre-Training

vs A100 on large language models.

NVIDIA B200 on Denvr AI Cloud

AI Ascend Assets-17.png

1T Parameter Training

Scale across 8-GPU NVLink nodes for 1,440 GB of total VRAM and 1.8 TB/s per-GPU interconnect.

AI Ascend Assets-17.png

LLM Inference

Serve 70B+ parameter models on a single GPU, or scale to 8 GPUs with 640 GB VRAM for 500B+ parameter models at FP8 precision.

AI Ascend Assets-17.png

Production AI

H100 delivers 9x faster training vs. A100 at $2.10/hr on-demand. For extended context or larger models, consider the H200 with 141 GB HBM3e.

AI Ascend Assets-17.png

Managed Storage

High-performance Weka filesystem and local NVMe available. No external storage to provision for datasets, checkpoints, or model artifacts.

Configurations

Per-minute billing with on-demand and reserved options. All configurations available as bare metal, VM, or model endpoints.

Platform

GPUs

On-Demand

VRAM

vCPUs

Memory

Local Storage

Interconnect

NVIDIA H100 SXM

8x

80 GB

208

1024 GB

6x 3.8TB NVMe

IB 3200G

$2.30 / GPU

NVIDIA H100 SXM

GPU COUNT

GPU VRAM

vCPUs

MEMORY

LOCAL STORAGE

INTERCONNECT

8x

80 GB

208

1024 GB

6x 3.8TB NVMe

IB 3200G

$2.30 / GPU

Related GPUs

Compare Denvr GPU options by workload and performance requirements.

Intel Gaudi 2

Optimized

For

View Details

Optimized

For

View Details

Optimized

For

VRAM

VRAM

VRAM

VRAM

Memory Bandwidth

VRAM

VRAM

VRAM

FP64/FP32

VRAM

VRAM

VRAM

FP16

VRAM

VRAM

VRAM

FP8

VRAM

VRAM

VRAM

NVLink

VRAM

VRAM

VRAM

On-Demand Pricing

VRAM

VRAM

VRAM

PLATFORM

VRAM

BANDWIDTH

FP64/FP32

FP16

FP8

NVLINK

GPU Count

GPU VRAM

vCPUs

Memory

Local Storage

Interconnect

GPU Count

PLATFORM

VRAM

BANDWIDTH

FP64/FP32

FP16

FP8

NVLINK

GPU Count

GPU VRAM

vCPUs

Memory

Local Storage

Interconnect

GPU Count

PLATFORM

VRAM

BANDWIDTH

FP64/FP32

FP16

FP8

NVLINK

GPU Count

GPU VRAM

vCPUs

Memory

Local Storage

Interconnect

GPU Count

Infrastructure you can trust at scale

As an NVIDIA Cloud Partner we build and operate AI clusters following NVIDIA Reference Architectures. Your models and data are supported via strict privacy safeguards and SOC 2 Type 2 security practices.

soc-150x150-1.webp
bottom of page