Skip to content

45 offerings from 18 providers across serverless products and GPU instances.

60 complete comparable prices; incomplete and additive prices remain visible but are not ranked as totals.

12 of 16 serverless offerings include a scale-to-zero operating mode.

Published configurations: 80 GB

Serverless prices

All serverless GPUs →
Baseten.co
H100 · 80 GB VRAM/GPU
H100 · Serverless · fixed
  • USD 0.1083/minute
    compute time
Component price — not a total
H100 standard
scale to zero · dedicated
1 minute increments · bills deploying, scaling up, scaling down, making predictions
Beam.cloud
1× H100 · 80 GB VRAM/GPU
H100 · Serverless · fixed
  • USD 0.0000125/core second
    cpu core time
  • USD 0.000972/second
    gpu time
  • USD 0.0000021/gib second
    memory time
Component price — not a total
Standard serverless
scale to zero
1 ms increments · bills code running
stale Verified 17 Sept 2026 Official price source (opens in new tab)
The Serverless section includes a “With committed spend” / “Book a call” option without a published amount; it was omitted.
Cerebrium
1× H100 · VRAM unknown
H100 · Serverless · fixed
  • USD 0.00000655/core second
    cpu core time
  • USD 0.000944/second
    gpu time
  • USD 0.00000222/gb second
    memory time
Component price — not a total
Standard scale-to-zero
scale to zero
1 second increments
Enterprise pricing is listed as Custom and was omitted.
1 geography evidence item
Cerebrium
1× H100 · VRAM unknown
H100 · Serverless · fixed
  • USD 0.00000655/core second
    cpu core time
  • USD 0.000944/second
    gpu time
  • USD 0.00000222/gb second
    memory time
Component price — not a total
Standard
scale to zero
1 second increments · bills active compute time
EU compute EU-constrained
Sweden
Enterprise pricing is listed as Custom and was omitted.
3 geography evidence items
Fal.ai
1× H100 · 80 GB VRAM/GPU
H100 · Serverless · fixed
  • USD 4.5/hour
    compute time
≈ EUR 3.963/hour
Always on
always on
1 second increments · bills setup, idle, running, draining, terminating
The source lists discounted “As low as” prices, but those custom deployment prices are omitted because the terms require sales contact.
Fal.ai
1× H100 · 80 GB VRAM/GPU
H100 · Serverless · fixed
  • USD 4.5/hour
    compute time
≈ EUR 3.963/hour
Scale to zero
scale to zero
1 second increments · bills setup, idle, running, draining, terminating
The source lists discounted “As low as” prices, but those custom deployment prices are omitted because the terms require sales contact.
Modal
1× H100 · VRAM unknown
Nvidia H100 SXM5 · Serverless · fixed
  • USD 0.0000131/core second → USD 0.00001507/core second effective
    cpu core time
  • USD 0.001097/second → USD 0.001262/second effective
    gpu time
  • USD 0.00000222/gib second → USD 0.000002553/gib second effective
    memory time
Component price — not a total
Broad region
scale to zero
none minimum · bills running
stale Verified 17 Sept 2026 Official price source (opens in new tab)
5 geography evidence items
Modal
1× H100 · VRAM unknown
Nvidia H100 SXM5 · Serverless · fixed
  • USD 0.0000131/core second → USD 0.00002293/core second effective
    cpu core time
  • USD 0.001097/second → USD 0.00192/second effective
    gpu time
  • USD 0.00000222/gib second → USD 0.000003885/gib second effective
    memory time
Component price — not a total
Narrow region
scale to zero
none minimum · bills running
stale Verified 17 Sept 2026 Official price source (opens in new tab)
5 geography evidence items
Modal
1× H100 · VRAM unknown
Nvidia H100 SXM5 · Serverless · fixed
  • USD 0.0000131/core second
    cpu core time
  • USD 0.001097/second
    gpu time
  • USD 0.00000222/gib second
    memory time
Component price — not a total
Standard
scale to zero
none minimum · bills running
stale Verified 17 Sept 2026 Official price source (opens in new tab)
5 geography evidence items
OVHcloud
1× H100 · 80 GB VRAM/GPU
h100-1-gpu · Serverless · fixed
  • USD 3.39/hour
    compute time
≈ EUR 2.985/hour
scale-to-zero
scale to zero · dedicated
1 minute increments · bills SCALING, RUNNING
Replicate
2× H100 · 80 GB VRAM/GPU
2x Nvidia H100 GPU · Serverless · fixed
  • USD 0.00305/second
    compute time
≈ EUR 9.67/hour
Public models
request metered · commitment
bills request execution time
stale Verified 17 Sept 2026 Official price source (opens in new tab)
Replicate
4× H100 · 80 GB VRAM/GPU
4x Nvidia H100 GPU · Serverless · fixed
  • USD 0.0061/second
    compute time
≈ EUR 19.34/hour
Public models
request metered · commitment
bills request execution time
stale Verified 17 Sept 2026 Official price source (opens in new tab)
Replicate
8× H100 · 80 GB VRAM/GPU
8x Nvidia H100 GPU · Serverless · fixed
  • USD 0.0122/second
    compute time
≈ EUR 38.68/hour
Public models
request metered · commitment
bills request execution time
stale Verified 17 Sept 2026 Official price source (opens in new tab)
Runpod
1× H100 · 80 GB VRAM/GPU
H100 · Serverless · fixed
  • USD 4.79/hour
    gpu time
≈ EUR 4.218/hour
Flex
scale to zero
1 second increments · bills initialization, execution, idle timeout
Active workers have discounts available through sales inquiry; no public price is published, so Active variants were omitted.

Instance prices

All GPU instances →
Verda
1× H100 SXM5 · 80 GB/GPU · 30 vCPU / 120 GB RAM
1× H100
80 GB VRAM/GPU · 30 vCPUs · 120 GB RAM · storage unknown
Region: unknown
  • USD 1.81/hour
    compute time
≈ EUR 1.594/hour
Spot
spot
Verda
1× H100 SXM5 · 80 GB/GPU · 32 vCPU / 170 GB RAM
1× H100
80 GB VRAM/GPU · 32 vCPUs · 170 GB RAM · storage unknown
Region: unknown
  • USD 1.81/hour
    compute time
≈ EUR 1.594/hour
Spot
spot
Nebius
NVIDIA H100 · 1 GPU / 16 vCPU / 200 GiB
1× H100
80 GB VRAM/GPU · 16 vCPUs · 200 GB RAM · storage unknown
Region: Finland · FI · eu-north1 · EU
  • USD 2.15/hour
    compute time
≈ EUR 1.893/hour
eu-north1 preemptible
spot
stale Verified 17 Sept 2026 Official price source (opens in new tab)
Verda
1× H100 SXM5 · 80 GB/GPU · 30 vCPU / 120 GB RAM
1× H100
80 GB VRAM/GPU · 30 vCPUs · 120 GB RAM · storage unknown
Region: unknown
  • USD 3.63/hour
    compute time
≈ EUR 3.197/hour
On-demand
standard
Verda
1× H100 SXM5 · 80 GB/GPU · 32 vCPU / 170 GB RAM
1× H100
80 GB VRAM/GPU · 32 vCPUs · 170 GB RAM · storage unknown
Region: unknown
  • USD 3.63/hour
    compute time
≈ EUR 3.197/hour
On-demand
standard
Nebius
NVIDIA H100 · 1 GPU / 16 vCPU / 200 GiB
1× H100
80 GB VRAM/GPU · 16 vCPUs · 200 GB RAM · storage unknown
Region: Finland · FI · eu-north1 · EU
  • USD 3.85/hour
    compute time
≈ EUR 3.391/hour
eu-north1 on-demand
standard
stale Verified 17 Sept 2026 Official price source (opens in new tab)
Lambda
2× NVIDIA H100 SXM
2× H100
80 GB VRAM/GPU · 52 vCPUs · 450 GB RAM · 5632 GB storage
Region: unknown
  • USD 4.19/hour → USD 8.38/hour effective
    compute time
≈ EUR 7.38/hour
On-demand
standard
Lambda
4× NVIDIA H100 SXM
4× H100
80 GB VRAM/GPU · 104 vCPUs · 900 GB RAM · 11264 GB storage
Region: unknown
  • USD 4.09/hour → USD 16.36/hour effective
    compute time
≈ EUR 14.41/hour
On-demand
standard
Nebius
NVIDIA H100 · 8 GPU / 128 vCPU / 1600 GiB
8× H100
80 GB VRAM/GPU · 128 vCPUs · 1600 GB RAM · storage unknown
Region: Finland · FI · eu-north1 · EU
  • USD 17.2/hour
    compute time
≈ EUR 15.15/hour
eu-north1 preemptible
spot
stale Verified 17 Sept 2026 Official price source (opens in new tab)
Nebius
NVIDIA H100 · 8 GPU / 128 vCPU / 1600 GiB
8× H100
80 GB VRAM/GPU · 128 vCPUs · 1600 GB RAM · storage unknown
Region: Finland · FI · eu-north1 · EU
  • USD 30.8/hour
    compute time
≈ EUR 27.12/hour
eu-north1 on-demand
standard
stale Verified 17 Sept 2026 Official price source (opens in new tab)
Lambda
8× NVIDIA H100 SXM
8× H100
80 GB VRAM/GPU · 208 vCPUs · 1800 GB RAM · 22528 GB storage
Region: unknown
  • USD 3.99/hour → USD 31.92/hour effective
    compute time
≈ EUR 28.11/hour
On-demand
standard
CoreWeave
NVIDIA HGX H100
8× H100
80 GB VRAM/GPU · 128 vCPUs · 2048 GB RAM · 62914.56 GB storage
Region: north-america
  • USD 49.24/hour
    compute time
≈ EUR 43.36/hour
North America on-demand
standard

Available Providers