This page is part of the FinOpsForge ontology — a structured library of named FinOps entities, each treated with consistent operations: define, implement, compare, calculate. Full GPU Price Index methodology →
What This Index Is (and Isn't)
This index normalizes published cloud GPU pricing to a single comparable unit — USD per GPU-hour — across two hyperscalers with full on-demand pricing (Azure, GCP), three GPU-specialist "neoclouds" (Lambda, CoreWeave, RunPod), and AWS EC2 Capacity Blocks for ML reservation pricing (see below). Multi-GPU instances (for example, an 8×H100 CoreWeave or GCP instance) are normalized by dividing the published instance price by the GPU count in that SKU; providers that already price per single GPU (Lambda, RunPod, GCP's T4 GPU-attachment SKU) are used as-is. Every price is a snapshot as of the "checked" date shown on each row — cloud GPU pricing changes frequently, so treat any single figure as a point-in-time reference and re-check the source link before budgeting against it.
AWS's regular EC2 on-demand and Spot GPU pricing is still not represented in this index: the standard EC2 pricing pages remain JavaScript-rendered, and the unauthenticated bulk pricing offer file is too large to retrieve in a single request (see the methodology page for detail). This edition does add 6 AWS EC2 Capacity Blocks for ML rows (checked 2026-07-20) — a distinct reservation-fee pricing model where you pre-schedule a future block of GPU capacity in advance, not an instant-launch on-demand rate. Capacity Blocks rows are labeled “Capacity Blocks” everywhere they appear in the tables below and are never blended with on-demand figures. Vast.ai is excluded on principle: it is a peer-to-peer marketplace with prices that vary by listing rather than a single vendor-published rate.
Cheapest Normalized $/GPU-hour by Model
The table below shows, for each GPU model covered by this index, the single cheapest normalized price found across all six providers (on-demand rates for every provider except AWS, whose entries reflect Capacity Blocks reservation pricing instead — see the “Price type” column). This is the citable, at-a-glance view — the full per-provider SKU breakdown (with spot and committed-use pricing) follows in the next section.
| GPU model | Cheapest normalized $/GPU-hr | Price type | Provider | SKU | Region | Checked |
|---|---|---|---|---|---|---|
| T4 | $0.35 | On-demand | GCP | T4 GPU attachment (1x) | us-central1 (Iowa) | 2026-07-20 |
| L4 | $0.39 | On-demand | RunPod | L4 (Pods, Secure Cloud) | Secure Cloud (multi-region, not region-differentiated on pricing page) | 2026-07-20 |
| V100 | $0.79 | On-demand | Lambda | 1x NVIDIA Tesla V100 (Instances) | US (single global price list, no region selector) | 2026-07-20 |
| L40S | $0.99 | On-demand | RunPod | L40S (Pods, Secure Cloud) | Secure Cloud (multi-region, not region-differentiated on pricing page) | 2026-07-20 |
| L40 | $1.25 | On-demand | CoreWeave | NVIDIA L40 (8x) | North America | 2026-07-20 |
| A100 | $1.39 | On-demand | RunPod | A100 PCIe (Pods, Secure Cloud) | Secure Cloud (multi-region, not region-differentiated on pricing page) | 2026-07-20 |
| H100 | $2.89 | On-demand | RunPod | H100 PCIe (Pods, Secure Cloud) | Secure Cloud (multi-region, not region-differentiated on pricing page) | 2026-07-20 |
| H200 | $4.39 | On-demand | RunPod | H200 SXM (Pods, Secure Cloud) | Secure Cloud (multi-region, not region-differentiated on pricing page) | 2026-07-20 |
| B200 | $5.89 | On-demand | RunPod | B200 (Pods, Secure Cloud) | Secure Cloud (multi-region, not region-differentiated on pricing page) | 2026-07-20 |
| GH200 | $6.50 | On-demand | CoreWeave | NVIDIA GH200 (1x) | North America | 2026-07-20 |
| MI300X | $7.20 | On-demand | Azure | Standard_ND96isr_MI300X_v5 (8x) | canadacentral | 2026-07-20 |
| B300 | $7.39 | On-demand | RunPod | B300 (Pods, Secure Cloud) | Secure Cloud (multi-region, not region-differentiated on pricing page) | 2026-07-20 |
Full Pricing by Provider
Every priced SKU in this index, grouped by provider, with on-demand, spot/preemptible (where publicly listed), and 1-year/3-year committed-use pricing, all normalized to USD per GPU-hour. Spot and preemptible prices are volatile point-in-time snapshots, not guaranteed rates — GCP, for example, states that Spot prices are dynamic and can change up to once every 30 days.
AWS
AWS rows below are EC2 Capacity Blocks for ML reservation-fee prices — you reserve a future block of GPU capacity in advance for a fixed duration, which is a different purchasing model from the instant-launch on-demand rates shown for every other provider in this table (see the “Price type” column below and the methodology page for the full explanation). Regular AWS on-demand and Spot GPU pricing is still not represented in this index.
| SKU | GPU | GPU count | VRAM (GB) | Region | Price type | Capacity Block $/GPU-hr | Spot $/GPU-hr | 1yr commit $/GPU-hr | 3yr commit $/GPU-hr | Source | Checked |
|---|---|---|---|---|---|---|---|---|---|---|---|
| p5.48xlarge (EC2 Capacity Blocks for ML) | H100 | 8 | 80 | US East (N. Virginia) | Capacity Blocks | $4.33 | — | — | — | source | 2026-07-20 |
| p5e.48xlarge (EC2 Capacity Blocks for ML) | H200 | 8 | 141 | US East (Ohio) | Capacity Blocks | $4.97 | — | — | — | source | 2026-07-20 |
| p5en.48xlarge (EC2 Capacity Blocks for ML) | H200 | 8 | 141 | US East (N. Virginia) | Capacity Blocks | $5.72 | — | — | — | source | 2026-07-20 |
| p4d.24xlarge (EC2 Capacity Blocks for ML) | A100 | 8 | 40 | US East (N. Virginia) | Capacity Blocks | $1.48 | — | — | — | source | 2026-07-20 |
| p6-b300.48xlarge (EC2 Capacity Blocks for ML) | B300 | 8 | 288 | US East (N. Virginia) | Capacity Blocks | $11.70 | — | — | — | source | 2026-07-20 |
| p6-b200.48xlarge (EC2 Capacity Blocks for ML) | B200 | 8 | 180 | US East (N. Virginia) | Capacity Blocks | $10.30 | — | — | — | source | 2026-07-20 |
Azure
| SKU | GPU | GPU count | VRAM (GB) | Region | Price type | On-demand $/GPU-hr | Spot $/GPU-hr | 1yr commit $/GPU-hr | 3yr commit $/GPU-hr | Source | Checked |
|---|---|---|---|---|---|---|---|---|---|---|---|
| Standard_ND96isr_H100_v5 (8x) | H100 | 8 | 80 | eastus | On-demand | $12.29 | $2.27 | $7.87 | $5.40 | source | 2026-07-20 |
| Standard_ND96asr_v4 (8x) | A100 | 8 | 40 | eastus | On-demand | $3.40 | $0.75 | $2.35 | $1.36 | source | 2026-07-20 |
| Standard_ND96isr_MI300X_v5 (8x) | MI300X | 8 | 192 | canadacentral | On-demand | $7.20 | $1.33 | $4.61 | $3.16 | source | 2026-07-20 |
CoreWeave
| SKU | GPU | GPU count | VRAM (GB) | Region | Price type | On-demand $/GPU-hr | Spot $/GPU-hr | 1yr commit $/GPU-hr | 3yr commit $/GPU-hr | Source | Checked |
|---|---|---|---|---|---|---|---|---|---|---|---|
| NVIDIA HGX B200 (8x) | B200 | 8 | 180 | North America | On-demand | $8.60 | $4.26 | — | — | source | 2026-07-20 |
| NVIDIA HGX H200 (8x) | H200 | 8 | 141 | North America | On-demand | $6.30 | $2.62 | — | — | source | 2026-07-20 |
| NVIDIA HGX H100 (8x) | H100 | 8 | 80 | North America | On-demand | $6.16 | $2.46 | — | — | source | 2026-07-20 |
| NVIDIA GH200 (1x) | GH200 | 1 | 96 | North America | On-demand | $6.50 | — | — | — | source | 2026-07-20 |
| NVIDIA A100 80GB (8x) | A100 | 8 | 80 | North America | On-demand | $2.70 | $1.21 | — | — | source | 2026-07-20 |
| NVIDIA L40S (8x) | L40S | 8 | 48 | North America | On-demand | $2.25 | $0.98 | — | — | source | 2026-07-20 |
| NVIDIA L40 (8x) | L40 | 8 | 48 | North America | On-demand | $1.25 | $0.78 | — | — | source | 2026-07-20 |
GCP
| SKU | GPU | GPU count | VRAM (GB) | Region | Price type | On-demand $/GPU-hr | Spot $/GPU-hr | 1yr commit $/GPU-hr | 3yr commit $/GPU-hr | Source | Checked |
|---|---|---|---|---|---|---|---|---|---|---|---|
| a3-highgpu-8g | H100 | 8 | 80 | us-central1 (Iowa) | On-demand | $11.06 | $4.74 | $7.67 | $4.86 | source | 2026-07-20 |
| a2-highgpu-8g | A100 | 8 | 40 | us-central1 (Iowa) | On-demand | $3.67 | $1.80 | $2.31 | $1.29 | source | 2026-07-20 |
| g2-standard-96 | L4 | 8 | 24 | us-central1 (Iowa) | On-demand | $1.00 | $0.44 | $0.63 | $0.45 | source | 2026-07-20 |
| T4 GPU attachment (1x) | T4 | 1 | 16 | us-central1 (Iowa) | On-demand | $0.35 | — | $0.22 | $0.16 | source | 2026-07-20 |
Lambda
| SKU | GPU | GPU count | VRAM (GB) | Region | Price type | On-demand $/GPU-hr | Spot $/GPU-hr | 1yr commit $/GPU-hr | 3yr commit $/GPU-hr | Source | Checked |
|---|---|---|---|---|---|---|---|---|---|---|---|
| 1x NVIDIA B200 SXM6 (Instances) | B200 | 1 | 180 | US (single global price list, no region selector) | On-demand | $6.69 | — | — | — | source | 2026-07-20 |
| 1x NVIDIA H100 SXM (Instances) | H100 | 1 | 80 | US (single global price list, no region selector) | On-demand | $3.99 | — | — | — | source | 2026-07-20 |
| 1x NVIDIA A100 SXM 80GB (Instances) | A100 | 1 | 80 | US (single global price list, no region selector) | On-demand | $2.79 | — | — | — | source | 2026-07-20 |
| 1x NVIDIA A100 SXM 40GB (Instances) | A100 | 1 | 40 | US (single global price list, no region selector) | On-demand | $1.99 | — | — | — | source | 2026-07-20 |
| 1x NVIDIA Tesla V100 (Instances) | V100 | 1 | 16 | US (single global price list, no region selector) | On-demand | $0.79 | — | — | — | source | 2026-07-20 |
RunPod
| SKU | GPU | GPU count | VRAM (GB) | Region | Price type | On-demand $/GPU-hr | Spot $/GPU-hr | 1yr commit $/GPU-hr | 3yr commit $/GPU-hr | Source | Checked |
|---|---|---|---|---|---|---|---|---|---|---|---|
| B300 (Pods, Secure Cloud) | B300 | 1 | 288 | Secure Cloud (multi-region, not region-differentiated on pricing page) | On-demand | $7.39 | — | — | — | source | 2026-07-20 |
| B200 (Pods, Secure Cloud) | B200 | 1 | 180 | Secure Cloud (multi-region, not region-differentiated on pricing page) | On-demand | $5.89 | — | — | — | source | 2026-07-20 |
| H200 SXM (Pods, Secure Cloud) | H200 | 1 | 141 | Secure Cloud (multi-region, not region-differentiated on pricing page) | On-demand | $4.39 | — | — | — | source | 2026-07-20 |
| H100 SXM (Pods, Secure Cloud) | H100 | 1 | 80 | Secure Cloud (multi-region, not region-differentiated on pricing page) | On-demand | $2.99 | — | — | — | source | 2026-07-20 |
| H100 NVL 94GB (Pods, Secure Cloud) | H100 | 1 | 94 | Secure Cloud (multi-region, not region-differentiated on pricing page) | On-demand | $3.19 | — | — | — | source | 2026-07-20 |
| H100 PCIe (Pods, Secure Cloud) | H100 | 1 | 80 | Secure Cloud (multi-region, not region-differentiated on pricing page) | On-demand | $2.89 | — | — | — | source | 2026-07-20 |
| A100 SXM (Pods, Secure Cloud) | A100 | 1 | 80 | Secure Cloud (multi-region, not region-differentiated on pricing page) | On-demand | $1.49 | — | — | — | source | 2026-07-20 |
| A100 PCIe (Pods, Secure Cloud) | A100 | 1 | 80 | Secure Cloud (multi-region, not region-differentiated on pricing page) | On-demand | $1.39 | — | — | — | source | 2026-07-20 |
| L40S (Pods, Secure Cloud) | L40S | 1 | 48 | Secure Cloud (multi-region, not region-differentiated on pricing page) | On-demand | $0.99 | — | — | — | source | 2026-07-20 |
| L4 (Pods, Secure Cloud) | L4 | 1 | 24 | Secure Cloud (multi-region, not region-differentiated on pricing page) | On-demand | $0.39 | — | — | — | source | 2026-07-20 |
Methodology Summary
Full detail — including region selection per provider, exactly how multi-GPU instances are normalized, what's excluded and why, and the complete source list — lives on the dedicated GPU Price Index methodology page. In short: one representative, named region per provider where applicable (Azure: eastus; GCP: us-central1/Iowa; CoreWeave: North America; Lambda and RunPod: single global/tier-wide price lists; AWS: no single representative region — EC2 Capacity Blocks pricing is quoted per SKU/region on AWS's own pricing page, so each AWS row states its own region individually); prices sourced only from each vendor's own pricing page or public API, never from third-party aggregators; and any SKU that could not be live-verified at build time is dropped from the index rather than estimated.
This index applies the same GPU-specific purchasing-model logic covered in GPU Cost Optimization and the same commitment/spot mechanics covered generally in AWS Spot Instances Guide — this page exists to make the actual current numbers behind those strategies checkable in one place. For how GPU spend fits the broader AI cost picture, see FinOps for AI.
// FAQ
Estimate your cloud savings
Free FinOps Savings Calculator — AWS, Azure & GCP · no signup