FinOpsForge — Independent cloud cost reviews. No vendor sponsorships. No paid rankings.

GPU Cloud Price Index 2026: AWS Capacity Blocks, Azure, GCP, Lambda, CoreWeave & RunPod

// FinOps Data Asset — index(GPU Cloud Pricing) // July 2026 // independently researched
// Affiliate disclosure: FinOpsForge may earn a commission if you sign up via links on this page. This never affects the prices published in this index. Every figure below is sourced directly from the listed provider's own pricing page or API, not from an affiliate feed.
The FinOpsForge GPU Cloud Price Index is a normalized, dated table of cloud GPU pricing — on-demand, spot/preemptible, and 1–3 year committed-use rates converted to a common USD-per-GPU-hour unit — collected directly from each provider's own published pricing page or public pricing API. No third-party pricing aggregators are used as a source, and no price is estimated or interpolated: every row that appears below has a live-verified source URL and a "checked" date, and any row that could not be independently verified at build time is dropped rather than guessed.
// Editorial Methodology
This page is part of the FinOpsForge ontology — a structured library of named FinOps entities, each treated with consistent operations: define, implement, compare, calculate. Full GPU Price Index methodology →

What This Index Is (and Isn't)

This index normalizes published cloud GPU pricing to a single comparable unit — USD per GPU-hour — across two hyperscalers with full on-demand pricing (Azure, GCP), three GPU-specialist "neoclouds" (Lambda, CoreWeave, RunPod), and AWS EC2 Capacity Blocks for ML reservation pricing (see below). Multi-GPU instances (for example, an 8×H100 CoreWeave or GCP instance) are normalized by dividing the published instance price by the GPU count in that SKU; providers that already price per single GPU (Lambda, RunPod, GCP's T4 GPU-attachment SKU) are used as-is. Every price is a snapshot as of the "checked" date shown on each row — cloud GPU pricing changes frequently, so treat any single figure as a point-in-time reference and re-check the source link before budgeting against it.

AWS's regular EC2 on-demand and Spot GPU pricing is still not represented in this index: the standard EC2 pricing pages remain JavaScript-rendered, and the unauthenticated bulk pricing offer file is too large to retrieve in a single request (see the methodology page for detail). This edition does add 6 AWS EC2 Capacity Blocks for ML rows (checked 2026-07-20) — a distinct reservation-fee pricing model where you pre-schedule a future block of GPU capacity in advance, not an instant-launch on-demand rate. Capacity Blocks rows are labeled “Capacity Blocks” everywhere they appear in the tables below and are never blended with on-demand figures. Vast.ai is excluded on principle: it is a peer-to-peer marketplace with prices that vary by listing rather than a single vendor-published rate.

Cheapest Normalized $/GPU-hour by Model

The table below shows, for each GPU model covered by this index, the single cheapest normalized price found across all six providers (on-demand rates for every provider except AWS, whose entries reflect Capacity Blocks reservation pricing instead — see the “Price type” column). This is the citable, at-a-glance view — the full per-provider SKU breakdown (with spot and committed-use pricing) follows in the next section.

GPU modelCheapest normalized $/GPU-hrPrice typeProviderSKURegionChecked
T4$0.35On-demandGCPT4 GPU attachment (1x)us-central1 (Iowa)2026-07-20
L4$0.39On-demandRunPodL4 (Pods, Secure Cloud)Secure Cloud (multi-region, not region-differentiated on pricing page)2026-07-20
V100$0.79On-demandLambda1x NVIDIA Tesla V100 (Instances)US (single global price list, no region selector)2026-07-20
L40S$0.99On-demandRunPodL40S (Pods, Secure Cloud)Secure Cloud (multi-region, not region-differentiated on pricing page)2026-07-20
L40$1.25On-demandCoreWeaveNVIDIA L40 (8x)North America2026-07-20
A100$1.39On-demandRunPodA100 PCIe (Pods, Secure Cloud)Secure Cloud (multi-region, not region-differentiated on pricing page)2026-07-20
H100$2.89On-demandRunPodH100 PCIe (Pods, Secure Cloud)Secure Cloud (multi-region, not region-differentiated on pricing page)2026-07-20
H200$4.39On-demandRunPodH200 SXM (Pods, Secure Cloud)Secure Cloud (multi-region, not region-differentiated on pricing page)2026-07-20
B200$5.89On-demandRunPodB200 (Pods, Secure Cloud)Secure Cloud (multi-region, not region-differentiated on pricing page)2026-07-20
GH200$6.50On-demandCoreWeaveNVIDIA GH200 (1x)North America2026-07-20
MI300X$7.20On-demandAzureStandard_ND96isr_MI300X_v5 (8x)canadacentral2026-07-20
B300$7.39On-demandRunPodB300 (Pods, Secure Cloud)Secure Cloud (multi-region, not region-differentiated on pricing page)2026-07-20

Full Pricing by Provider

Every priced SKU in this index, grouped by provider, with on-demand, spot/preemptible (where publicly listed), and 1-year/3-year committed-use pricing, all normalized to USD per GPU-hour. Spot and preemptible prices are volatile point-in-time snapshots, not guaranteed rates — GCP, for example, states that Spot prices are dynamic and can change up to once every 30 days.

AWS

AWS rows below are EC2 Capacity Blocks for ML reservation-fee prices — you reserve a future block of GPU capacity in advance for a fixed duration, which is a different purchasing model from the instant-launch on-demand rates shown for every other provider in this table (see the “Price type” column below and the methodology page for the full explanation). Regular AWS on-demand and Spot GPU pricing is still not represented in this index.

SKUGPUGPU countVRAM (GB)RegionPrice typeCapacity Block $/GPU-hrSpot $/GPU-hr1yr commit $/GPU-hr3yr commit $/GPU-hrSourceChecked
p5.48xlarge (EC2 Capacity Blocks for ML)H100880US East (N. Virginia)Capacity Blocks$4.33source2026-07-20
p5e.48xlarge (EC2 Capacity Blocks for ML)H2008141US East (Ohio)Capacity Blocks$4.97source2026-07-20
p5en.48xlarge (EC2 Capacity Blocks for ML)H2008141US East (N. Virginia)Capacity Blocks$5.72source2026-07-20
p4d.24xlarge (EC2 Capacity Blocks for ML)A100840US East (N. Virginia)Capacity Blocks$1.48source2026-07-20
p6-b300.48xlarge (EC2 Capacity Blocks for ML)B3008288US East (N. Virginia)Capacity Blocks$11.70source2026-07-20
p6-b200.48xlarge (EC2 Capacity Blocks for ML)B2008180US East (N. Virginia)Capacity Blocks$10.30source2026-07-20

Azure

SKUGPUGPU countVRAM (GB)RegionPrice typeOn-demand $/GPU-hrSpot $/GPU-hr1yr commit $/GPU-hr3yr commit $/GPU-hrSourceChecked
Standard_ND96isr_H100_v5 (8x)H100880eastusOn-demand$12.29$2.27$7.87$5.40source2026-07-20
Standard_ND96asr_v4 (8x)A100840eastusOn-demand$3.40$0.75$2.35$1.36source2026-07-20
Standard_ND96isr_MI300X_v5 (8x)MI300X8192canadacentralOn-demand$7.20$1.33$4.61$3.16source2026-07-20

CoreWeave

SKUGPUGPU countVRAM (GB)RegionPrice typeOn-demand $/GPU-hrSpot $/GPU-hr1yr commit $/GPU-hr3yr commit $/GPU-hrSourceChecked
NVIDIA HGX B200 (8x)B2008180North AmericaOn-demand$8.60$4.26source2026-07-20
NVIDIA HGX H200 (8x)H2008141North AmericaOn-demand$6.30$2.62source2026-07-20
NVIDIA HGX H100 (8x)H100880North AmericaOn-demand$6.16$2.46source2026-07-20
NVIDIA GH200 (1x)GH200196North AmericaOn-demand$6.50source2026-07-20
NVIDIA A100 80GB (8x)A100880North AmericaOn-demand$2.70$1.21source2026-07-20
NVIDIA L40S (8x)L40S848North AmericaOn-demand$2.25$0.98source2026-07-20
NVIDIA L40 (8x)L40848North AmericaOn-demand$1.25$0.78source2026-07-20

GCP

SKUGPUGPU countVRAM (GB)RegionPrice typeOn-demand $/GPU-hrSpot $/GPU-hr1yr commit $/GPU-hr3yr commit $/GPU-hrSourceChecked
a3-highgpu-8gH100880us-central1 (Iowa)On-demand$11.06$4.74$7.67$4.86source2026-07-20
a2-highgpu-8gA100840us-central1 (Iowa)On-demand$3.67$1.80$2.31$1.29source2026-07-20
g2-standard-96L4824us-central1 (Iowa)On-demand$1.00$0.44$0.63$0.45source2026-07-20
T4 GPU attachment (1x)T4116us-central1 (Iowa)On-demand$0.35$0.22$0.16source2026-07-20

Lambda

SKUGPUGPU countVRAM (GB)RegionPrice typeOn-demand $/GPU-hrSpot $/GPU-hr1yr commit $/GPU-hr3yr commit $/GPU-hrSourceChecked
1x NVIDIA B200 SXM6 (Instances)B2001180US (single global price list, no region selector)On-demand$6.69source2026-07-20
1x NVIDIA H100 SXM (Instances)H100180US (single global price list, no region selector)On-demand$3.99source2026-07-20
1x NVIDIA A100 SXM 80GB (Instances)A100180US (single global price list, no region selector)On-demand$2.79source2026-07-20
1x NVIDIA A100 SXM 40GB (Instances)A100140US (single global price list, no region selector)On-demand$1.99source2026-07-20
1x NVIDIA Tesla V100 (Instances)V100116US (single global price list, no region selector)On-demand$0.79source2026-07-20

RunPod

SKUGPUGPU countVRAM (GB)RegionPrice typeOn-demand $/GPU-hrSpot $/GPU-hr1yr commit $/GPU-hr3yr commit $/GPU-hrSourceChecked
B300 (Pods, Secure Cloud)B3001288Secure Cloud (multi-region, not region-differentiated on pricing page)On-demand$7.39source2026-07-20
B200 (Pods, Secure Cloud)B2001180Secure Cloud (multi-region, not region-differentiated on pricing page)On-demand$5.89source2026-07-20
H200 SXM (Pods, Secure Cloud)H2001141Secure Cloud (multi-region, not region-differentiated on pricing page)On-demand$4.39source2026-07-20
H100 SXM (Pods, Secure Cloud)H100180Secure Cloud (multi-region, not region-differentiated on pricing page)On-demand$2.99source2026-07-20
H100 NVL 94GB (Pods, Secure Cloud)H100194Secure Cloud (multi-region, not region-differentiated on pricing page)On-demand$3.19source2026-07-20
H100 PCIe (Pods, Secure Cloud)H100180Secure Cloud (multi-region, not region-differentiated on pricing page)On-demand$2.89source2026-07-20
A100 SXM (Pods, Secure Cloud)A100180Secure Cloud (multi-region, not region-differentiated on pricing page)On-demand$1.49source2026-07-20
A100 PCIe (Pods, Secure Cloud)A100180Secure Cloud (multi-region, not region-differentiated on pricing page)On-demand$1.39source2026-07-20
L40S (Pods, Secure Cloud)L40S148Secure Cloud (multi-region, not region-differentiated on pricing page)On-demand$0.99source2026-07-20
L4 (Pods, Secure Cloud)L4124Secure Cloud (multi-region, not region-differentiated on pricing page)On-demand$0.39source2026-07-20

Methodology Summary

Full detail — including region selection per provider, exactly how multi-GPU instances are normalized, what's excluded and why, and the complete source list — lives on the dedicated GPU Price Index methodology page. In short: one representative, named region per provider where applicable (Azure: eastus; GCP: us-central1/Iowa; CoreWeave: North America; Lambda and RunPod: single global/tier-wide price lists; AWS: no single representative region — EC2 Capacity Blocks pricing is quoted per SKU/region on AWS's own pricing page, so each AWS row states its own region individually); prices sourced only from each vendor's own pricing page or public API, never from third-party aggregators; and any SKU that could not be live-verified at build time is dropped from the index rather than estimated.

This index applies the same GPU-specific purchasing-model logic covered in GPU Cost Optimization and the same commitment/spot mechanics covered generally in AWS Spot Instances Guide — this page exists to make the actual current numbers behind those strategies checkable in one place. For how GPU spend fits the broader AI cost picture, see FinOps for AI.

// FAQ

Which cloud has the cheapest GPU pricing?
It depends entirely on the GPU model and purchasing mode you need — there is no single cheapest provider across the board. Use the "Cheapest Normalized $/GPU-hour by Model" table above to compare the lowest on-demand rate per GPU model, then check the full pricing table for spot and committed-use options on the specific provider you're evaluating.
Is AWS included in this index?
Partially. AWS EC2 Capacity Blocks for ML reservation pricing (6 SKUs, checked 2026-07-20) is included and clearly labeled “Capacity Blocks” in the Price type column — this is a reservation fee for a pre-scheduled future block of GPU capacity, not an instant-launch rate. Regular AWS EC2 on-demand and Spot GPU pricing is still not included: those pricing pages are JavaScript-rendered and the bulk pricing offer file is too large to fetch in one request. See the methodology page for full detail.
How often is this index updated?
The index is refreshed on a monthly cadence by re-running the same data collection and generator process against each provider's live pricing source. The "Last updated" date at the top of this page reflects the most recent refresh.
Why are some rows missing spot or committed-use pricing?
Not every provider publishes every purchasing mode for every SKU. Lambda and RunPod's Pods pricing publishes on-demand rates only (both list committed capacity as "contact sales" with no public number); CoreWeave publishes on-demand and spot but only a qualitative "up to 60% off, contact sales" claim for reserved capacity, so no committed-use number is shown for CoreWeave. Fields are left blank rather than estimated in every case.
🧮

Estimate your cloud savings

Free FinOps Savings Calculator — AWS, Azure & GCP · no signup

Try it free →

Estimate Your Cloud Savings

Free calculator — no signup required. AWS, Azure & GCP supported.

Try the FinOps Savings Calculator →

// Related