Skip to content

Cloud GPU Pricing Catalog

Compare GPU prices across providers. Every price links to its official source and shows when it was last verified.

Results

199 matching offerings · 331 billing options · filtered

Baseten.co
NVIDIA B200 · 180 GB VRAM/GPU
B200 · Serverless · fixed
  • USD 0.1663/minute
    compute time
Component price — not a total
Standard
scale to zero · dedicated
1 minute increments · bills deploying, scaling up, scaling down, making predictions
Baseten.co
H100 · 80 GB VRAM/GPU
H100 · Serverless · fixed
  • USD 0.1083/minute
    compute time
Component price — not a total
Standard
scale to zero · dedicated
1 minute increments · bills deploying, scaling up, scaling down, making predictions
Baseten.co
A100 · 80 GB VRAM/GPU
A100 · Serverless · fixed
  • USD 0.06667/minute
    compute time
Component price — not a total
Standard
scale to zero · dedicated
1 minute increments · bills deploying, scaling up, scaling down, making predictions
Beam.cloud
1× NVIDIA B200 · 180 GB VRAM/GPU
B200 · Serverless · fixed
  • USD 0.0000125/core second
    cpu core time
  • USD 0.001561/second
    gpu time
  • USD 0.0000021/gib second
    memory time
Component price — not a total
Standard
scale to zero
1 ms increments · bills code running
The Serverless committed-spend option is presented as “With committed spend Book a call →” without a published amount; it was omitted.
Beam.cloud
1× L40S · 48 GB VRAM/GPU
L40S · Serverless · fixed
  • USD 0.0000125/core second
    cpu core time
  • USD 0.000486/second
    gpu time
  • USD 0.0000021/gib second
    memory time
Component price — not a total
Standard
scale to zero
1 ms increments · bills code running
The Serverless committed-spend option is presented as “With committed spend Book a call →” without a published amount; it was omitted.
Beam.cloud
1× A100 · 80 GB VRAM/GPU
A100 80GB · Serverless · fixed
  • USD 0.0000125/core second
    cpu core time
  • USD 0.000625/second
    gpu time
  • USD 0.0000021/gib second
    memory time
Component price — not a total
Standard
scale to zero
1 ms increments · bills code running
The Serverless committed-spend option is presented as “With committed spend Book a call →” without a published amount; it was omitted.
Beam.cloud
1× H100 · 80 GB VRAM/GPU
H100 · Serverless · fixed
  • USD 0.0000125/core second
    cpu core time
  • USD 0.000986/second
    gpu time
  • USD 0.0000021/gib second
    memory time
Component price — not a total
Standard
scale to zero
1 ms increments · bills code running
The Serverless committed-spend option is presented as “With committed spend Book a call →” without a published amount; it was omitted.
Beam.cloud
1× NVIDIA RTX PRO 6000 Blackwell · 96 GB VRAM/GPU
RTX PRO 6000 · Serverless · fixed
  • USD 0.0000125/core second
    cpu core time
  • USD 0.000758/second
    gpu time
  • USD 0.0000021/gib second
    memory time
Component price — not a total
Standard
scale to zero
1 ms increments · bills code running
The Serverless committed-spend option is presented as “With committed spend Book a call →” without a published amount; it was omitted.
Beam.cloud
1× NVIDIA H200 · 141 GB VRAM/GPU
H200 · Serverless · fixed
  • USD 0.0000125/core second
    cpu core time
  • USD 0.001136/second
    gpu time
  • USD 0.0000021/gib second
    memory time
Component price — not a total
Standard
scale to zero
1 ms increments · bills code running
The Serverless committed-spend option is presented as “With committed spend Book a call →” without a published amount; it was omitted.
Beam.cloud
1× A6000 · 48 GB VRAM/GPU
A6000 · Serverless · fixed
  • USD 0.0000125/core second
    cpu core time
  • USD 0.0002278/second
    gpu time
  • USD 0.0000021/gib second
    memory time
Component price — not a total
Standard
scale to zero
1 ms increments · bills code running
The Serverless committed-spend option is presented as “With committed spend Book a call →” without a published amount; it was omitted.
Cerebrium
1× A100 · 80 GB VRAM/GPU
A100 (80GB) · Serverless · fixed
  • USD 0.00000655/core second
    cpu core time
  • USD 0.000583/second
    gpu time
  • USD 0.00000222/gb second
    memory time
Component price — not a total
Standard
scale to zero
1 second increments · bills actual compute time
EU compute EU-constrained
Sweden
3 geography evidence items
CoreWeave
8× NVIDIA H200 · 141 GB VRAM/GPU
NVIDIA HGX H200 · GPU VM · fixed
  • USD 50.44/hour
    compute time
≈ EUR 43.92/hour
Europe on-demand
instance rental
bills running
EU compute
Spain
3 geography evidence items
CoreWeave
8× NVIDIA H200 · 141 GB VRAM/GPU
NVIDIA HGX H200 · GPU VM · fixed
  • USD 20.64/hour
    compute time
≈ EUR 17.97/hour
Europe spot
instance rental · spot
bills running
EU compute
Spain
3 geography evidence items
CoreWeave
8× NVIDIA H200 · 141 GB VRAM/GPU
NVIDIA HGX H200 · GPU VM · fixed
  • USD 50.44/hour
    compute time
≈ EUR 43.92/hour
North America on-demand
instance rental
bills running
EU compute
Spain
3 geography evidence items
CoreWeave
8× NVIDIA H200 · 141 GB VRAM/GPU
NVIDIA HGX H200 · GPU VM · fixed
  • USD 20.93/hour
    compute time
≈ EUR 18.22/hour
North America spot
instance rental · spot
bills running
EU compute
Spain
3 geography evidence items
CoreWeave
8× NVIDIA RTX PRO 6000 Blackwell · 96 GB VRAM/GPU
NVIDIA RTX PRO 6000 Blackwell Server Edition · GPU VM · fixed
  • USD 20/hour
    compute time
≈ EUR 17.41/hour
Europe on-demand
instance rental
bills running
EU compute
Spain
3 geography evidence items
CoreWeave
8× NVIDIA RTX PRO 6000 Blackwell · 96 GB VRAM/GPU
NVIDIA RTX PRO 6000 Blackwell Server Edition · GPU VM · fixed
  • USD 11.01/hour
    compute time
≈ EUR 9.586/hour
Europe spot
instance rental · spot
bills running
EU compute
Spain
3 geography evidence items
CoreWeave
8× NVIDIA RTX PRO 6000 Blackwell · 96 GB VRAM/GPU
NVIDIA RTX PRO 6000 Blackwell Server Edition · GPU VM · fixed
  • USD 20/hour
    compute time
≈ EUR 17.41/hour
North America on-demand
instance rental
bills running
EU compute
Spain
3 geography evidence items
CoreWeave
8× NVIDIA RTX PRO 6000 Blackwell · 96 GB VRAM/GPU
NVIDIA RTX PRO 6000 Blackwell Server Edition · GPU VM · fixed
  • USD 11.09/hour
    compute time
≈ EUR 9.656/hour
North America spot
instance rental · spot
bills running
EU compute
Spain
3 geography evidence items
CoreWeave
8× NVIDIA B200 · 180 GB VRAM/GPU
NVIDIA HGX B200 · GPU VM · fixed
  • USD 68.8/hour
    compute time
≈ EUR 59.9/hour
Europe on-demand
instance rental
bills running
EU compute
Spain
3 geography evidence items
CoreWeave
8× NVIDIA B200 · 180 GB VRAM/GPU
NVIDIA HGX B200 · GPU VM · fixed
  • USD 34.87/hour
    compute time
≈ EUR 30.36/hour
Europe spot
instance rental · spot
bills running
EU compute
Spain
3 geography evidence items
CoreWeave
8× NVIDIA B200 · 180 GB VRAM/GPU
NVIDIA HGX B200 · GPU VM · fixed
  • USD 68.8/hour
    compute time
≈ EUR 59.9/hour
North America on-demand
instance rental
bills running
EU compute
Spain
3 geography evidence items
CoreWeave
8× NVIDIA B200 · 180 GB VRAM/GPU
NVIDIA HGX B200 · GPU VM · fixed
  • USD 34.11/hour
    compute time
≈ EUR 29.7/hour
North America spot
instance rental · spot
bills running
EU compute
Spain
3 geography evidence items
DigitalOcean
1× AMD Instinct MI300X · 192 GB VRAM/GPU
AMD MI300X · GPU VM · fixed
  • USD 1.99/hour
    compute time
≈ EUR 1.733/hour
On-demand
instance rental
1 second increments · 1 minute minimum · bills provisioned
DigitalOcean
8× H100 · 80 GB VRAM/GPU
NVIDIA H100 · 8 GPU · GPU VM · fixed
  • USD 23.92/hour
    compute time
≈ EUR 20.83/hour
On-demand
instance rental
1 second increments · 1 minute minimum · bills provisioned
DigitalOcean
8× NVIDIA H200 · 141 GB VRAM/GPU
NVIDIA H200 · 8 GPU · GPU VM · fixed
  • USD 27.52/hour
    compute time
≈ EUR 23.96/hour
On-demand
instance rental
1 second increments · 1 minute minimum · bills provisioned
DigitalOcean
8× AMD Instinct MI300X · 192 GB VRAM/GPU
AMD MI300X · 8 GPU · GPU VM · fixed
  • USD 15.92/hour
    compute time
≈ EUR 13.86/hour
On-demand
instance rental
1 second increments · 1 minute minimum · bills provisioned
DigitalOcean
1× NVIDIA H200 · 141 GB VRAM/GPU
NVIDIA H200 · GPU VM · fixed
  • USD 3.44/hour
    compute time
≈ EUR 2.995/hour
On-demand
instance rental
1 second increments · 1 minute minimum · bills provisioned
DigitalOcean
1× RTX 6000 Ada · 48 GB VRAM/GPU
NVIDIA RTX 6000 · GPU VM · fixed
  • USD 1.57/hour
    compute time
≈ EUR 1.367/hour
On-demand
instance rental
1 second increments · 1 minute minimum · bills provisioned
Fal.ai
1× NVIDIA B200 · 180 GB VRAM/GPU
B200 · Serverless · fixed
  • USD 6.25/hour
    compute time
≈ EUR 5.442/hour
B200 always-on
always on
1 second increments · bills setup, idle, running, draining, terminating
The Serverless & Compute Pricing table publishes As low as amounts, but those discounted prices require sales contact and are omitted.
Fal.ai
1× NVIDIA B200 · 180 GB VRAM/GPU
B200 · Serverless · fixed
  • USD 6.25/hour
    compute time
≈ EUR 5.442/hour
B200 scale-to-zero
scale to zero
1 second increments · bills setup, idle, running, draining, terminating
The Serverless & Compute Pricing table publishes As low as amounts, but those discounted prices require sales contact and are omitted.
Fal.ai
1× NVIDIA B300 · 288 GB VRAM/GPU
B300 · Serverless · fixed
  • USD 8.5/hour
    compute time
≈ EUR 7.401/hour
B300 always-on
always on
1 second increments · bills setup, idle, running, draining, terminating
The Serverless & Compute Pricing table publishes As low as amounts, but those discounted prices require sales contact and are omitted.
Fal.ai
1× NVIDIA B300 · 288 GB VRAM/GPU
B300 · Serverless · fixed
  • USD 8.5/hour
    compute time
≈ EUR 7.401/hour
B300 scale-to-zero
scale to zero
1 second increments · bills setup, idle, running, draining, terminating
The Serverless & Compute Pricing table publishes As low as amounts, but those discounted prices require sales contact and are omitted.
Fal.ai
1× H100 · 80 GB VRAM/GPU
H100 · Serverless · fixed
  • USD 3.99/hour
    compute time
≈ EUR 3.474/hour
H100 always-on
always on
1 second increments · bills setup, idle, running, draining, terminating
The Serverless & Compute Pricing table publishes As low as amounts, but those discounted prices require sales contact and are omitted.
Fal.ai
1× H100 · 80 GB VRAM/GPU
H100 · Serverless · fixed
  • USD 3.99/hour
    compute time
≈ EUR 3.474/hour
H100 scale-to-zero
scale to zero
1 second increments · bills setup, idle, running, draining, terminating
The Serverless & Compute Pricing table publishes As low as amounts, but those discounted prices require sales contact and are omitted.
Fal.ai
1× NVIDIA RTX PRO 6000 Blackwell · 96 GB VRAM/GPU
RTX PRO 6000 · Serverless · fixed
  • USD 2.99/hour
    compute time
≈ EUR 2.603/hour
RTX PRO 6000 always-on
always on
1 second increments · bills setup, idle, running, draining, terminating
The Serverless & Compute Pricing table publishes As low as amounts, but those discounted prices require sales contact and are omitted.
Fal.ai
1× NVIDIA RTX PRO 6000 Blackwell · 96 GB VRAM/GPU
RTX PRO 6000 · Serverless · fixed
  • USD 2.99/hour
    compute time
≈ EUR 2.603/hour
RTX PRO 6000 scale-to-zero
scale to zero
1 second increments · bills setup, idle, running, draining, terminating
The Serverless & Compute Pricing table publishes As low as amounts, but those discounted prices require sales contact and are omitted.
Fal.ai
1× NVIDIA H200 · 141 GB VRAM/GPU
H200 · Serverless · fixed
  • USD 4.5/hour
    compute time
≈ EUR 3.918/hour
H200 always-on
always on
1 second increments · bills setup, idle, running, draining, terminating
The Serverless & Compute Pricing table publishes As low as amounts, but those discounted prices require sales contact and are omitted.
Fal.ai
1× NVIDIA H200 · 141 GB VRAM/GPU
H200 · Serverless · fixed
  • USD 4.5/hour
    compute time
≈ EUR 3.918/hour
H200 scale-to-zero
scale to zero
1 second increments · bills setup, idle, running, draining, terminating
The Serverless & Compute Pricing table publishes As low as amounts, but those discounted prices require sales contact and are omitted.
Fireworks AI
1× NVIDIA B200 · 180 GB VRAM/GPU
B200 180 GB GPU · Dedicated GPU · fixed
  • USD 10/hour
    gpu time
≈ EUR 8.707/hour
Always on
always on · dedicated
1 second increments · bills active GPU replicas
Fireworks AI
1× NVIDIA B200 · 180 GB VRAM/GPU
B200 180 GB GPU · Dedicated GPU · fixed
  • USD 10/hour
    gpu time
≈ EUR 8.707/hour
Scale to zero
scale to zero · dedicated
1 second increments · bills active GPU replicas
Fireworks AI
1× NVIDIA H200 · 141 GB VRAM/GPU
H200 141 GB GPU · Dedicated GPU · fixed
  • USD 7/hour
    gpu time
≈ EUR 6.095/hour
Scale to zero
scale to zero · dedicated
1 second increments · bills active GPU replicas
Fireworks AI
1× NVIDIA B300 · 288 GB VRAM/GPU
B300 288 GB GPU · Dedicated GPU · fixed
  • USD 12/hour
    gpu time
≈ EUR 10.45/hour
Always on
always on · dedicated
1 second increments · bills active GPU replicas
Fireworks AI
1× NVIDIA B300 · 288 GB VRAM/GPU
B300 288 GB GPU · Dedicated GPU · fixed
  • USD 12/hour
    gpu time
≈ EUR 10.45/hour
Scale to zero
scale to zero · dedicated
1 second increments · bills active GPU replicas
Google Cloud
1× NVIDIA RTX PRO 6000 Blackwell · 96 GB VRAM/GPU
NVIDIA RTX Pro 6000 · Serverless · fixed
  • USD 0.000018/core second
    cpu core time
  • USD 0.0003652/second
    gpu time
  • USD 0.000002/gib second
    memory time
Component price — not a total
NVIDIA RTX Pro 6000 No zonal redundancy
scale to zero
100 ms increments · 1 minute minimum · bills entire instance lifecycle
EU compute EU-constrained
Netherlands
3 geography evidence items
Google Cloud
1× NVIDIA RTX PRO 6000 Blackwell · 96 GB VRAM/GPU
NVIDIA RTX Pro 6000 · Serverless · fixed
  • USD 0.000018/core second
    cpu core time
  • USD 0.0005691/second
    gpu time
  • USD 0.000002/gib second
    memory time
Component price — not a total
NVIDIA RTX Pro 6000 Zonal redundancy
scale to zero
100 ms increments · 1 minute minimum · bills entire instance lifecycle
EU compute EU-constrained
Netherlands
3 geography evidence items
Hugging Face
4× A100 · 80 GB VRAM/GPU
GCP NVIDIA A100 x4 · Dedicated GPU · fixed
  • USD 14.4/hour
    compute time
≈ EUR 12.54/hour
Always on
always on · dedicated
1 minute increments · 1 minute minimum · bills initializing, running
GPU rows do not publish system memory, so systemMemoryGb remains null.
Hugging Face
4× A100 · 80 GB VRAM/GPU
GCP NVIDIA A100 x4 · Dedicated GPU · fixed
  • USD 14.4/hour
    compute time
≈ EUR 12.54/hour
Scale to zero
scale to zero · dedicated
1 minute increments · 1 minute minimum · bills initializing, running
GPU rows do not publish system memory, so systemMemoryGb remains null.
Hugging Face
8× A100 · 80 GB VRAM/GPU
AWS NVIDIA A100 x8 · Dedicated GPU · fixed
  • USD 20/hour
    compute time
≈ EUR 17.41/hour
Always on
always on · dedicated
1 minute increments · 1 minute minimum · bills initializing, running
GPU rows do not publish system memory, so systemMemoryGb remains null.
Hugging Face
8× A100 · 80 GB VRAM/GPU
AWS NVIDIA A100 x8 · Dedicated GPU · fixed
  • USD 20/hour
    compute time
≈ EUR 17.41/hour
Scale to zero
scale to zero · dedicated
1 minute increments · 1 minute minimum · bills initializing, running
GPU rows do not publish system memory, so systemMemoryGb remains null.