Skip to content

59 offerings from 17 providers across serverless products and GPU instances.

80 complete comparable prices; incomplete and additive prices remain visible but are not ranked as totals.

19 of 23 serverless offerings include a scale-to-zero operating mode.

Published configurations: 40 GB · 80 GB

Serverless prices

All serverless GPUs →
Microsoft Azure
1× A100 · 80 GB VRAM/GPU
NVIDIA A100 · Serverless · fixed
  • USD 0.000024/core second
    cpu core time
  • USD 0.000529/second
    gpu time
  • USD 0.000003/gib second
    memory time
Component price — not a total
NVIDIA A100 — East US
scale to zero
1 second increments · bills active usage
EU compute EU-constrained
Italy · Sweden
5 geography evidence items
Microsoft Azure
1× A100 · 80 GB VRAM/GPU
NVIDIA A100 · Serverless · fixed
  • USD 0.000024/core second
    cpu core time
  • USD 0.000688/second
    gpu time
  • USD 0.000003/gib second
    memory time
Component price — not a total
NVIDIA A100 — Sweden Central
scale to zero
1 second increments · bills active usage
EU compute EU-constrained
Italy · Sweden
5 geography evidence items
Baseten.co
A100 · 80 GB VRAM/GPU
A100 · Serverless · fixed
  • USD 0.06667/minute
    compute time
Component price — not a total
Standard
scale to zero · dedicated
1 minute increments · bills deploying, scaling up, scaling down, making predictions
Beam.cloud
1× A100 · 80 GB VRAM/GPU
A100 80GB · Serverless · fixed
  • USD 0.0000125/core second
    cpu core time
  • USD 0.000625/second
    gpu time
  • USD 0.0000021/gib second
    memory time
Component price — not a total
Standard
scale to zero
1 ms increments · bills code running
The Serverless committed-spend option is presented as “With committed spend Book a call →” without a published amount; it was omitted.
Cerebrium
1× A100 · 40 GB VRAM/GPU
A100 (40GB) · Serverless · fixed
  • USD 0.00000655/core second
    cpu core time
  • USD 0.000555/second
    gpu time
  • USD 0.00000222/gb second
    memory time
Component price — not a total
Standard
scale to zero
1 second increments · bills actual compute time
EU compute EU-constrained
Sweden
3 geography evidence items
Cerebrium
1× A100 · 80 GB VRAM/GPU
A100 (80GB) · Serverless · fixed
  • USD 0.00000655/core second
    cpu core time
  • USD 0.000583/second
    gpu time
  • USD 0.00000222/gb second
    memory time
Component price — not a total
Standard
scale to zero
1 second increments · bills actual compute time
EU compute EU-constrained
Sweden
3 geography evidence items
Inferless
1× A100 · 80 GB VRAM/GPU
Nvidia A100 Dedicated · Serverless · fixed
  • USD 0.001491/second
    compute time
≈ EUR 4.674/hour
Standard
scale to zero · dedicated
1 second increments · bills model loading, healthy state, request processing
Enterprise pricing is listed as discounted or custom and has no published price; omitted from the catalog.
Inferless
1× A100 · 40 GB VRAM/GPU
Nvidia A100 Shared · Serverless · fixed
  • USD 0.000745/second
    compute time
≈ EUR 2.335/hour
Standard
scale to zero
1 second increments · bills model loading, healthy state, request processing
Enterprise pricing is listed as discounted or custom and has no published price; omitted from the catalog.
Modal
1× A100 · 40 GB VRAM/GPU
Nvidia A100, 40 GB · Serverless · fixed
  • USD 0.0000131/core second → USD 0.00001965/core second effective
    cpu core time
  • USD 0.000583/second → USD 0.0008745/second effective
    gpu time
  • USD 0.00000222/gib second → USD 0.00000333/gib second effective
    memory time
Component price — not a total
Broad region
scale to zero
none minimum · bills actual compute time
5 geography evidence items
Modal
1× A100 · 40 GB VRAM/GPU
Nvidia A100, 40 GB · Serverless · fixed
  • USD 0.0000131/core second → USD 0.00002293/core second effective
    cpu core time
  • USD 0.000583/second → USD 0.00102/second effective
    gpu time
  • USD 0.00000222/gib second → USD 0.000003885/gib second effective
    memory time
Component price — not a total
Narrow region
scale to zero
none minimum · bills actual compute time
5 geography evidence items
Modal
1× A100 · 40 GB VRAM/GPU
Nvidia A100, 40 GB · Serverless · fixed
  • USD 0.0000131/core second
    cpu core time
  • USD 0.000583/second
    gpu time
  • USD 0.00000222/gib second
    memory time
Component price — not a total
Standard region
scale to zero
none minimum · bills actual compute time
5 geography evidence items
Modal
1× A100 · 80 GB VRAM/GPU
Nvidia A100, 80 GB · Serverless · fixed
  • USD 0.0000131/core second → USD 0.00001965/core second effective
    cpu core time
  • USD 0.000694/second → USD 0.001041/second effective
    gpu time
  • USD 0.00000222/gib second → USD 0.00000333/gib second effective
    memory time
Component price — not a total
Broad region
scale to zero
none minimum · bills actual compute time
5 geography evidence items
Modal
1× A100 · 80 GB VRAM/GPU
Nvidia A100, 80 GB · Serverless · fixed
  • USD 0.0000131/core second → USD 0.00002293/core second effective
    cpu core time
  • USD 0.000694/second → USD 0.001215/second effective
    gpu time
  • USD 0.00000222/gib second → USD 0.000003885/gib second effective
    memory time
Component price — not a total
Narrow region
scale to zero
none minimum · bills actual compute time
5 geography evidence items
Modal
1× A100 · 80 GB VRAM/GPU
Nvidia A100, 80 GB · Serverless · fixed
  • USD 0.0000131/core second
    cpu core time
  • USD 0.000694/second
    gpu time
  • USD 0.00000222/gib second
    memory time
Component price — not a total
Standard region
scale to zero
none minimum · bills actual compute time
5 geography evidence items
OVHcloud
1× A100 · 80 GB VRAM/GPU
a100-1-gpu · Serverless · fixed
  • USD 3.35/hour
    compute time
≈ EUR 2.917/hour
a100-1-gpu always-on
always on · dedicated
1 minute increments · bills SCALING, RUNNING
OVHcloud
1× A100 · 80 GB VRAM/GPU
a100-1-gpu · Serverless · fixed
  • USD 3.35/hour
    compute time
≈ EUR 2.917/hour
a100-1-gpu scale-to-zero
scale to zero · dedicated
1 minute increments · bills SCALING, RUNNING
Replicate
4× A100 · 80 GB VRAM/GPU
4x Nvidia A100 (80GB) GPU · Serverless · fixed
  • USD 0.0056/second
    compute time
≈ EUR 17.55/hour
Committed multi-GPU capacity
request metered · commitment
bills request execution time
Replicate
8× A100 · 80 GB VRAM/GPU
8x Nvidia A100 (80GB) GPU · Serverless · fixed
  • USD 0.0112/second
    compute time
≈ EUR 35.11/hour
Committed multi-GPU capacity
request metered · commitment
bills request execution time
Runpod
1× A100 · 80 GB VRAM/GPU
A100 · Serverless · fixed
  • USD 2.72/hour
    gpu time
≈ EUR 2.368/hour
Flex
scale to zero
1 second increments · bills initialization, execution, idle timeout
Active worker discounts are available through sales inquiry and are not published; custom pricing plans are available through sales inquiry, so no Active or custom-priced variants were included.
Runpod offers custom pricing plans for large scale and enterprise workloads; contact sales for pricing.

Instance prices

All GPU instances →
Verda
1× A100 SXM4 · 40 GB/GPU
1× A100
40 GB VRAM/GPU · 22 vCPUs · 120 GB RAM · storage unknown
Region: unknown
  • USD 0.4515/hour
    compute time
≈ EUR 0.3931/hour
Spot
spot
Verda
1× A100 SXM4 · 80 GB/GPU
1× A100
80 GB VRAM/GPU · 22 vCPUs · 120 GB RAM · storage unknown
Region: unknown
  • USD 0.6265/hour
    compute time
≈ EUR 0.5455/hour
Spot
spot
Verda
1× A100 SXM4 · 40 GB/GPU
1× A100
40 GB VRAM/GPU · 22 vCPUs · 120 GB RAM · storage unknown
Region: unknown
  • USD 1.29/hour
    compute time
≈ EUR 1.123/hour
On-demand
standard
Verda
1× A100 SXM4 · 80 GB/GPU
1× A100
80 GB VRAM/GPU · 22 vCPUs · 120 GB RAM · storage unknown
Region: unknown
  • USD 1.79/hour
    compute time
≈ EUR 1.559/hour
On-demand
standard
Paperspace
A100-80G
1× A100
80 GB VRAM/GPU · 12 vCPUs · 90 GB RAM · storage unknown
Region: New York City, New York · US · ny2
  • USD 3.18/hour
    compute time
≈ EUR 2.769/hour
NY2 on-demand
standard
Microsoft Azure
Standard_NC24ads_A100_v4
1× A100
80 GB VRAM/GPU · 24 vCPUs · 220 GB RAM · 1024 GB storage
Region: East US (Virginia) · US · eastus
  • USD 3.673/hour
    compute time
≈ EUR 3.198/hour
EASTUS Linux pay-as-you-go
standard
Google Cloud
a2-highgpu-1g
1× A100
40 GB VRAM/GPU · 12 vCPUs · 85 GB RAM · storage unknown
Region: Council Bluffs, Iowa · US · us-central1
  • USD 3.673/hour
    compute time
≈ EUR 3.198/hour
US-CENTRAL1 on-demand
standard
Google Cloud
a2-highgpu-1g
1× A100
40 GB VRAM/GPU · 12 vCPUs · 85 GB RAM · storage unknown
Region: Eemshaven, Netherlands · NL · europe-west4 · EU
  • USD 3.748/hour
    compute time
≈ EUR 3.263/hour
EUROPE-WEST4 on-demand
standard
Lambda
2× NVIDIA A100 PCIe
2× A100
40 GB VRAM/GPU · 60 vCPUs · 450 GB RAM · 1024 GB storage
Region: unknown
  • USD 1.99/hour → USD 3.98/hour effective
    compute time
≈ EUR 3.465/hour
On-demand
standard
Microsoft Azure
Standard_NC24ads_A100_v4
1× A100
80 GB VRAM/GPU · 24 vCPUs · 220 GB RAM · 1024 GB storage
Region: West Europe (Netherlands) · NL · westeurope · EU
  • USD 4.775/hour
    compute time
≈ EUR 4.158/hour
WESTEUROPE Linux pay-as-you-go
standard
Google Cloud
a2-highgpu-2g
2× A100
40 GB VRAM/GPU · 24 vCPUs · 170 GB RAM · storage unknown
Region: Council Bluffs, Iowa · US · us-central1
  • USD 7.347/hour
    compute time
≈ EUR 6.397/hour
US-CENTRAL1 on-demand
standard
Google Cloud
a2-highgpu-2g
2× A100
40 GB VRAM/GPU · 24 vCPUs · 170 GB RAM · storage unknown
Region: Eemshaven, Netherlands · NL · europe-west4 · EU
  • USD 7.496/hour
    compute time
≈ EUR 6.527/hour
EUROPE-WEST4 on-demand
standard
Lambda
4× NVIDIA A100 PCIe
4× A100
40 GB VRAM/GPU · 120 vCPUs · 900 GB RAM · 1024 GB storage
Region: unknown
  • USD 1.99/hour → USD 7.96/hour effective
    compute time
≈ EUR 6.931/hour
On-demand
standard
Google Cloud
a2-highgpu-4g
4× A100
40 GB VRAM/GPU · 48 vCPUs · 340 GB RAM · storage unknown
Region: Council Bluffs, Iowa · US · us-central1
  • USD 14.69/hour
    compute time
≈ EUR 12.79/hour
US-CENTRAL1 on-demand
standard
Google Cloud
a2-highgpu-4g
4× A100
40 GB VRAM/GPU · 48 vCPUs · 340 GB RAM · storage unknown
Region: Eemshaven, Netherlands · NL · europe-west4 · EU
  • USD 14.99/hour
    compute time
≈ EUR 13.05/hour
EUROPE-WEST4 on-demand
standard
Lambda
8× NVIDIA A100 SXM · 40 GB/GPU
8× A100
40 GB VRAM/GPU · 124 vCPUs · 1800 GB RAM · 5939.2 GB storage
Region: unknown
  • USD 1.99/hour → USD 15.92/hour effective
    compute time
≈ EUR 13.86/hour
On-demand
standard
CoreWeave
NVIDIA A100
8× A100
80 GB VRAM/GPU · 128 vCPUs · 2048 GB RAM · 7864.32 GB storage
Region: north-america
  • USD 21.6/hour
    compute time
≈ EUR 18.81/hour
North America on-demand
standard
Lambda
8× NVIDIA A100 SXM · 80 GB/GPU
8× A100
80 GB VRAM/GPU · 240 vCPUs · 1800 GB RAM · 19968 GB storage
Region: unknown
  • USD 2.79/hour → USD 22.32/hour effective
    compute time
≈ EUR 19.43/hour
On-demand
standard
Google Cloud
a2-highgpu-8g
8× A100
40 GB VRAM/GPU · 96 vCPUs · 680 GB RAM · storage unknown
Region: Council Bluffs, Iowa · US · us-central1
  • USD 29.39/hour
    compute time
≈ EUR 25.59/hour
US-CENTRAL1 on-demand
standard
Google Cloud
a2-highgpu-8g
8× A100
40 GB VRAM/GPU · 96 vCPUs · 680 GB RAM · storage unknown
Region: Eemshaven, Netherlands · NL · europe-west4 · EU
  • USD 29.98/hour
    compute time
≈ EUR 26.11/hour
EUROPE-WEST4 on-demand
standard
Google Cloud
a2-megagpu-16g
16× A100
40 GB VRAM/GPU · 96 vCPUs · 1360 GB RAM · storage unknown
Region: Council Bluffs, Iowa · US · us-central1
  • USD 55.74/hour
    compute time
≈ EUR 48.53/hour
US-CENTRAL1 on-demand
standard
Google Cloud
a2-megagpu-16g
16× A100
40 GB VRAM/GPU · 96 vCPUs · 1360 GB RAM · storage unknown
Region: Eemshaven, Netherlands · NL · europe-west4 · EU
  • USD 56.63/hour
    compute time
≈ EUR 49.3/hour
EUROPE-WEST4 on-demand
standard

Available Providers