Skip to content

45 offerings from 18 providers across serverless products and GPU instances.

61 complete comparable prices; incomplete and additive prices remain visible but are not ranked as totals.

11 of 15 serverless offerings include a scale-to-zero operating mode.

Published configurations: 80 GB · 160 GB · 320 GB · 640 GB

Serverless prices

All serverless GPUs →
Baseten.co
H100 · 80 GB VRAM
H100 · Serverless · fixed
  • USD 0.1083/minute
    compute time
Component price — not a total
Standard
scale to zero · dedicated
1 minute increments · bills deploying, scaling up, scaling down, making predictions
Beam.cloud
1× H100 · 80 GB VRAM
H100 · Serverless · fixed
  • USD 0.0000125/core second
    cpu core time
  • USD 0.000986/second
    gpu time
  • USD 0.0000021/gib second
    memory time
Component price — not a total
Standard
scale to zero
1 ms increments · bills code running
The Serverless committed-spend option is presented as “With committed spend Book a call →” without a published amount; it was omitted.
Cerebrium
1× H100 · VRAM unknown
H100 · Serverless · fixed
  • USD 0.00000655/core second
    cpu core time
  • USD 0.000944/second
    gpu time
  • USD 0.00000222/gb second
    memory time
Component price — not a total
Standard
scale to zero
1 second increments · bills actual compute time
EU compute EU-constrained
2 geography evidence items
Fal.ai
1× H100 · 80 GB VRAM
H100 · Serverless · fixed
  • USD 3.99/hour
    compute time
≈ EUR 3.489/hour
H100 always-on
always on
1 second increments · bills setup, idle, running, draining, terminating
The Serverless & Compute Pricing table publishes As low as amounts, but those discounted prices require sales contact and are omitted.
Fal.ai
1× H100 · 80 GB VRAM
H100 · Serverless · fixed
  • USD 3.99/hour
    compute time
≈ EUR 3.489/hour
H100 scale-to-zero
scale to zero
1 second increments · bills setup, idle, running, draining, terminating
The Serverless & Compute Pricing table publishes As low as amounts, but those discounted prices require sales contact and are omitted.
Modal
1× H100 · VRAM unknown
Nvidia H100 · Serverless · fixed
  • USD 0.0000131/core second → USD 0.00001965/core second effective
    cpu core time
  • USD 0.001097/second → USD 0.001646/second effective
    gpu time
  • USD 0.00000222/gib second → USD 0.00000333/gib second effective
    memory time
Component price — not a total
Broad region
scale to zero
none minimum · bills actual compute time
5 geography evidence items
Modal
1× H100 · VRAM unknown
Nvidia H100 · Serverless · fixed
  • USD 0.0000131/core second → USD 0.00002293/core second effective
    cpu core time
  • USD 0.001097/second → USD 0.00192/second effective
    gpu time
  • USD 0.00000222/gib second → USD 0.000003885/gib second effective
    memory time
Component price — not a total
Narrow region
scale to zero
none minimum · bills actual compute time
5 geography evidence items
Modal
1× H100 · VRAM unknown
Nvidia H100 · Serverless · fixed
  • USD 0.0000131/core second
    cpu core time
  • USD 0.001097/second
    gpu time
  • USD 0.00000222/gib second
    memory time
Component price — not a total
Standard region
scale to zero
none minimum · bills actual compute time
5 geography evidence items
OVHcloud
1× H100 · 80 GB VRAM
h100-1-gpu · Serverless · fixed
  • USD 3.39/hour
    compute time
≈ EUR 2.965/hour
h100-1-gpu always-on
always on · dedicated
1 minute increments · bills SCALING, RUNNING
OVHcloud
1× H100 · 80 GB VRAM
h100-1-gpu · Serverless · fixed
  • USD 3.39/hour
    compute time
≈ EUR 2.965/hour
h100-1-gpu scale-to-zero
scale to zero · dedicated
1 minute increments · bills SCALING, RUNNING
Replicate
2× H100 · VRAM unknown
2x Nvidia H100 GPU · Serverless · fixed
  • USD 0.00305/second
    compute time
≈ EUR 9.602/hour
Committed multi-GPU capacity
request metered · commitment
bills request execution time
Replicate
4× H100 · VRAM unknown
4x Nvidia H100 GPU · Serverless · fixed
  • USD 0.0061/second
    compute time
≈ EUR 19.2/hour
Committed multi-GPU capacity
request metered · commitment
bills request execution time
Replicate
8× H100 · VRAM unknown
8x Nvidia H100 GPU · Serverless · fixed
  • USD 0.0122/second
    compute time
≈ EUR 38.41/hour
Committed multi-GPU capacity
request metered · commitment
bills request execution time
Runpod
1× H100 · 80 GB VRAM
H100 · Serverless · fixed
  • USD 4.55/hour
    gpu time
≈ EUR 3.979/hour
Flex
scale to zero
1 second increments · bills initialization, execution, idle timeout
Active worker discounts are available through sales inquiry and are omitted because no public price is published.

Instance prices

All GPU instances →
Nebius
NVIDIA H100 · 1 GPU / 16 vCPU / 200 GiB
1× H100
80 GB VRAM · 16 vCPUs · 200 GB RAM · storage unknown
Region: Finland · FI · eu-north1 · EU
  • USD 2.15/hour
    compute time
≈ EUR 1.88/hour
eu-north1 preemptible
spot
Nebius
NVIDIA H100 · 1 GPU / 16 vCPU / 200 GiB
1× H100
80 GB VRAM · 16 vCPUs · 200 GB RAM · storage unknown
Region: Finland · FI · eu-north1 · EU
  • USD 3.85/hour
    compute time
≈ EUR 3.367/hour
eu-north1 on-demand
standard
Lambda
2× NVIDIA H100 SXM
2× H100
80 GB VRAM · 52 vCPUs · 450 GB RAM · 5632 GB storage
Region: unknown
  • USD 4.19/hour → USD 8.38/hour effective
    compute time
≈ EUR 7.328/hour
On-demand
standard
Lambda
4× NVIDIA H100 SXM
4× H100
80 GB VRAM · 104 vCPUs · 900 GB RAM · 11264 GB storage
Region: unknown
  • USD 4.09/hour → USD 16.36/hour effective
    compute time
≈ EUR 14.31/hour
On-demand
standard
Nebius
NVIDIA H100 · 8 GPU / 128 vCPU / 1600 GiB
8× H100
80 GB VRAM · 128 vCPUs · 1600 GB RAM · storage unknown
Region: Finland · FI · eu-north1 · EU
  • USD 17.2/hour
    compute time
≈ EUR 15.04/hour
eu-north1 preemptible
spot
Nebius
NVIDIA H100 · 8 GPU / 128 vCPU / 1600 GiB
8× H100
80 GB VRAM · 128 vCPUs · 1600 GB RAM · storage unknown
Region: Finland · FI · eu-north1 · EU
  • USD 30.8/hour
    compute time
≈ EUR 26.93/hour
eu-north1 on-demand
standard
Lambda
8× NVIDIA H100 SXM
8× H100
80 GB VRAM · 208 vCPUs · 1800 GB RAM · 22528 GB storage
Region: unknown
  • USD 3.99/hour → USD 31.92/hour effective
    compute time
≈ EUR 27.91/hour
On-demand
standard

Available Providers