hostfleet Find my setup
data / live table

GPU cloud pricing

Public hourly list-rate equivalents for renting GPUs across 21 serverless, managed, and VM-style providers. This table is a rate reference, not a performance ranking or a deployment recommendation: GPU memory, billing lifecycle, fixed resources, region, and availability can change the practical choice. Per-second and per-minute rates are converted to hourly equivalents.

last verified: 2026-09-24 refresh cadence: Mon + Thu download raw JSON
GPU ↕ VRAM RunPod Pods ↕ RunPod Serverless ↕ Modal ↕ Fal.ai ↕ Baseten ↕ Replicate ↕ Lambda Cloud ↕ CoreWeave Inference ↕ Nebius AI Cloud ↕ Hyperstack ↕ Novita AI ↕ Verda (formerly DataCrunch) ↕ Salad Container Engine ↕ Vultr Cloud GPU ↕ DigitalOcean GPU Droplets ↕ Paperspace Machines ↕ Thunder Compute ↕ Massed Compute ↕ Jarvis Labs ↕ Koyeb ↕ Northflank ↕
M4000 8 GB ——————————————— $0.45/hr* —————
P4000 8 GB ——————————————— $0.51/hr* —————
RTX 4000 8 GB ——————————————— $0.56/hr* —————
P5000 16 GB ——————————————— $0.78/hr* —————
RTX 5000 16 GB ——————————————— $0.82/hr* —————
T4 16 GB —— $0.59/hr* — $0.63/hr* $0.81/hr* ———————————————
V100 16 GB 16 GB ——————————————— $2.30/hr* —————
RTX A4000 16 GB ————————— $0.15/hr* ————— $0.76/hr* —————
RTX 4000 Ada 20 GB —————————————— $0.76/hr* ———— $0.50/hr* —
L4 24 GB $0.49/hr* $0.69/hr* $0.80/hr* — $0.85/hr* ————————————— $0.44/hr* $0.70/hr* $0.80/hr*
A10 / A10G 24 GB —— $1.10/hr* — $1.21/hr* — $1.29/hr* ——————————————
Quadro RTX 6000 24 GB —————— $0.69/hr* ——————————————
P6000 24 GB ——————————————— $1.10/hr* —————
RTX A5000 24 GB ——————————————— $1.38/hr* — $0.44/hr* ———
A30 24 GB ————————————————— $0.35/hr* $0.41/hr* ——
RTX 4090 24 GB $0.74/hr* $1.10/hr ———————— $0.33/hr* — $0.16/hr* ————————
RTX 5090 32 GB $0.99/hr $1.58/hr ———————— $0.56/hr* — $0.25/hr* ————————
RTX PRO 4500 32 GB ————————————————— $0.92/hr* ———
V100 32 GB 32 GB ——————————————— $2.30/hr* —————
RTX A6000 / A40 48 GB $0.49/hr* $1.22/hr* ———— $1.09/hr* —— $0.50/hr* — $0.67/hr* — $1.71/hr* — $1.89/hr* $0.35/hr* $0.55/hr* — $0.75/hr* —
L40 48 GB ———————————————— $0.79/hr* $0.86/hr* ———
L40S 48 GB $1.09/hr* $1.75/hr* $1.95/hr* —— $3.51/hr* — $2.25/hr* $1.55/hr* — $0.55/hr* $1.56/hr* — $1.67/hr* $1.57/hr* —— $0.97/hr* — $1.20/hr* —
RTX 6000 Ada 48 GB —————————————— $1.57/hr* —— $0.79/hr* ———
RTX PRO 6000 96 GB $2.09/hr* $3.49/hr $3.03/hr* $2.99/hr* ——— $2.50/hr* $1.80/hr* $1.85/hr* — $2.06/hr* ————— $2.19/hr* $1.89/hr* $2.20/hr* $3.00/hr*
A100 40 GB 40 GB —— $2.10/hr* ——— $1.99/hr* ———— $1.31/hr* ——— $3.09/hr* —— $0.89/hr* — $1.42/hr*
A100 80 GB 80 GB $1.59/hr* $2.72/hr $2.50/hr* — $4.00/hr* $5.04/hr* — $2.70/hr* — $1.35/hr* — $1.82/hr* — $2.40/hr* — $3.18/hr* $1.09/hr* $1.35/hr* $1.49/hr* $1.60/hr* $1.76/hr*
H100 80 GB 80 GB $2.89/hr* $4.79/hr* $3.95/hr* $4.50/hr* $6.50/hr* $5.49/hr* $3.29/hr* $6.16/hr* $4.50/hr* $2.50/hr* $3.39/hr* $3.74/hr* —— $4.41/hr* $5.95/hr* $3.20/hr* $2.73/hr* $2.69/hr* $2.50/hr* $2.74/hr*
GH200 96 GB —————— $2.29/hr* ——————————————
H200 141 GB $4.59/hr* $5.93/hr $4.54/hr* $4.50/hr* ——— $6.31/hr* $5.40/hr* $3.99/hr* — $4.88/hr* —— $4.47/hr* —— $3.62/hr* $3.99/hr* $3.00/hr* —
B200 180 GB $6.79/hr* $8.64/hr $6.25/hr* $6.25/hr* $9.98/hr* — $6.99/hr* $8.60/hr* $8.50/hr* $6.00/hr* — $7.09/hr* ——————— $5.50/hr* —
AMD Instinct MI300X 192 GB —————————————— $2.59/hr* ——————
AMD Instinct MI325X 256 GB —————————————— $3.80/hr* ——————
B300 288 GB $7.89/hr* $9.98/hr $7.10/hr* $8.50/hr* ———— $9.50/hr* $7.40/hr* — $8.88/hr* —————————

Green = lowest displayed list rate for that GPU, not a recommendation. * = hover for the vendor's raw per-second/per-minute rate or a caveat. — = the vendor does not list that GPU publicly.

Sign up (affiliate, same price — some include credit): RunPod +$5 credit · Vultr +$300 trial · DigitalOcean GPU — header links above stay direct, always.

// billing models differ — read this before comparing

  • RunPod Pods — per-second, always-on container (Secure Cloud)
  • RunPod Serverless — per-second, scale-to-zero workers
  • Modal — per-second, scale-to-zero ($30/mo free credit)
  • Fal.ai — per-hour list price for custom deployments; committed-use discounts advertised
  • Baseten — per-minute, dedicated deployments, scale-to-zero
  • Replicate — per-second, private model deployments
  • Lambda Cloud — per-hour on-demand GPU instances; prices vary by instance GPU count
  • CoreWeave Inference — per-hour single-GPU inference rate; inference platform customers only (contact account executive)
  • Nebius AI Cloud — per-second VM compute, shown as on-demand GPU-hour; H100/H200/B200/B300/RTX PRO rates include prescribed vCPU and RAM, while L40S is the minimum all-in 1-GPU preset
  • Hyperstack — per-minute on-demand GPU VM, shown per GPU-hour; fixed CPU, RAM, root disk, and ephemeral disk are included, while public IPs and shared storage are separate
  • Novita AI — per-second on-demand GPU instance, settled hourly; listed one-GPU compute prices include the product's fixed vCPU and RAM, with storage above the free container-disk quota billed separately
  • Verda (formerly DataCrunch) — prepaid pay-as-you-go GPU instances in 10-minute increments, with unused terminated time refunded in the next billing period; listed one-GPU price includes fixed CPU and RAM, while storage is separate
  • Salad Container Engine — per-second managed container instances at Lowest (formerly Batch) priority; GPU rates include selected vCPU and RAM, allocation and image-download time are unbilled, and distributed nodes are interruptible
  • Vultr Cloud GPU — hourly on-demand GPU instances using actual calendar-month hours; listed instance rates include fixed vCPU, RAM, local storage, and outbound bandwidth allocations, and stopped instances continue billing until destroyed
  • DigitalOcean GPU Droplets — per-second on-demand GPU VMs with a 60-second or $0.01 minimum; listed one-GPU rates include fixed vCPU, RAM, boot/scratch storage, and outbound transfer allocations, and powered-off Droplets continue billing until destroyed
  • Paperspace Machines — per-hour GPU VM compute rates; compute billing stops only when Off, while the default 50 GB disk is separately billed at $0.0074/hour (maximum $5/month); dynamic IPs disappear at power-off, static IPs and other add-ons continue billing until removed; ingress and egress bandwidth are free
  • Thunder Compute — per-minute on-demand GPU VMs while running, shown as one-GPU base rates; four vCPUs and 8 GB RAM per vCPU are included in the base, additional vCPUs cost $0.04/hour, the first 100 GB persistent disk per GPU is included, and snapshots are separate
  • Massed Compute — per-minute debit for active on-demand GPU VMs using published hourly VM totals; one-GPU rates include fixed vCPU, RAM, and listed storage allocations; no bandwidth charges or long-term contracts
  • Jarvis Labs — per-minute on-demand managed Templates or GPU VMs, shown per GPU-hour; public one-GPU rows include listed vCPU and RAM, paused instances stop compute billing while storage continues at $0.10/GB-month, and in-region egress is free
  • Koyeb — per-second pay-per-use serverless GPU instances, shown at exact public hourly list rates; one-GPU configurations include fixed vCPU, RAM, and local disk, autoscaling is available, and saving-plan discounts are excluded
  • Northflank — per-second managed-cloud GPU workloads, shown at exact public GPU-hour component rates; CPU, memory, persistent disk, and egress are priced separately, GPU deployment requires at least $50 account credit, and model availability varies by region

// methodology

Prices are read from each vendor's public pricing page and converted to hourly equivalents (per-second × 3600, per-minute × 60), rounded to cents. "As low as" and committed-use rates are noted but never used in the comparison cells. We verify the whole table twice a week and stamp the date above; the raw JSON with source URLs is downloadable. If a number is wrong, the fix is one verification away — tell us.

Want cost per workload instead of cost per hour? Use the LLM hosting cost calculator.

Need to choose a deployment shape, not just a rate? Compare serverless GPU idle tails and scale-to-zero rules before treating a VM, serverless worker, and managed endpoint as interchangeable.

// signing up? these links support HostFleet

The source links in the table always go to official pricing pages, never through us. The signup links below are affiliate links — clearly labeled, same price for you, and some include free credit: