beam-logo
← All posts
Engineering

NVIDIA H100 Pricing

Eli MernitEli Mernit
September 29, 20263 min read
NVIDIA H100 Pricing

Renting an NVIDIA H100 PCIe costs $1.83 to $3.29 per GPU-hour on demand at specialist clouds. Hyperscaler VMs run $6.98 to $11.06 per GPU, and the top figure is Google's 8-GPU SXM VM. The median across 40 providers tracked by getdeploying is about $3.40, counting every H100 variant. Beam's on-demand H100 PCIe machine costs $1.83 per hour.

Key takeaways

  • Beam's $1.83 machine is the cheapest on-demand H100 PCIe here. Vast.ai ($1.93) and RunPod ($1.99) are within $0.16.
  • Google's on-demand H100 is an 8-GPU SXM VM at $88.49 per hour, so its per-GPU rate assumes you use all eight.
  • Interruptible capacity is far cheaper: Vast.ai's interruptible H100 PCIe offers run around $0.35 per GPU-hour.

H100 PCIe cloud prices per hour, by provider

ProviderSKUOn-demand $/GPU-hrNotes
BeamH100 PCIe$1.83On-demand machine; vCPU, RAM, and NVMe included
Vast.aiH100 PCIe (marketplace)$1.93Host-set price; median offer $2.57; low availability
RunPodH100 PCIe (Community Cloud)$1.99Per-second billing; Secure Cloud $2.89
HyperstackH100 PCIe$2.50Per-minute billing; spot $2.00; out of stock (getdeploying)
SpheronH100 PCIe$2.64Live marketplace rate; per-minute billing, 20-min minimum
LambdaH100 PCIe$3.29Fixed public rate; 1x instance only
AzureNC40ads H100 v5 (H100 NVL)$6.981-GPU VM; NVL is a 94 GB PCIe card; East US
Google Clouda3-highgpu-8g (8x H100 SXM)$11.068-GPU VM only ($88.49/hr), us-central1; SXM, not PCIe; spot $52.96/hr
Market median—$3.40Across 40 providers with an on-demand price (getdeploying); all H100 variants

How these prices were collected

Each figure is an on-demand list price per GPU-hour from the provider's own pricing page, with spot excluded. We used single-GPU shapes where offered and divided 8-GPU VM prices by eight.

H100 PCIe vs SXM5: what you trade away

The standard PCIe card is the slower H100: fewer SMs, HBM2e instead of HBM3, a lower power limit, and an NVLink bridge to at most one other card, where SXM5 boards join eight GPUs over NVSwitch. It keeps the Transformer Engine and FP8, so single-GPU inference and LoRA fine-tuning run well on it. The gap shows in multi-GPU distributed training, where SXM5 nodes win. Azure's H100 NVL sits between the two, a PCIe card with 94 GB.

Where to rent an H100 PCIe for less

On demand, nobody here undercuts Beam's $1.83, which includes vCPU, RAM, and NVMe. Vast.ai and RunPod come close, but Vast prices are set by individual hosts and change hourly, and RunPod's $1.99 is its Community Cloud; the Secure Cloud tier costs $2.89. Lambda charges $3.29 and bundles 26 vCPUs and 225 GiB of RAM. Paying less than $1.83 means interruptible capacity: Vast.ai's interruptible offers run around $0.35 per GPU-hour and can stop at any time, so checkpoint often.

What this means for your workload

An H100 PCIe suits single-GPU jobs that fit in 80 GB, like serving a mid-size LLM, LoRA or QLoRA fine-tuning, and FP8 batch inference. If weights plus KV cache outgrow 80 GB, an H200 with 141 GB beats splitting the model across PCIe cards; see our H200 pricing comparison.

FAQ

How much does an H100 PCIe cost per hour?

Specialist clouds charge $1.83 to $3.29 per GPU-hour on demand, with Beam at the low end and Lambda at the top. Azure's single-GPU H100 NVL VM costs $6.98 per hour.

What's the cheapest way to rent an H100 PCIe?

Beam's $1.83 on-demand machine is the lowest price we found, ahead of Vast.ai ($1.93) and RunPod Community Cloud ($1.99). Jobs that tolerate interruptions can use Vast.ai's interruptible offers at around $0.35 per GPU-hour.

H100 PCIe vs H100 SXM: which should I rent?

Rent the PCIe card for single-GPU inference and fine-tuning, where it costs less and keeps FP8 support. Rent SXM for multi-GPU training, because SXM boards link all eight GPUs through NVSwitch.

Eli Mernit
Eli Mernit
Published September 29, 2026
Pay as you gobilled by the millisecond

Start shipping on infra
you won’t outgrow.

Run sandboxes and GPU workloads on your cloud, and scale out to ours when you need to. No infra to manage.