Affiliate disclosure: HostFleet may earn a commission if you sign up through links on this page. That never changes the analysis. Read the live HostFleet about page for methodology and affiliate-policy context.

Source-backed rate-card and lifecycle comparison; calculated totals are estimates. This September refresh uses official provider pages, vendor APIs, and HostFleet’s live GPU pricing dataset. Every listed A100 rate was rechecked on September 14, 2026; Paperspace’s disk, IP, Machine-state, and Linux auto-shutdown rules were checked on September 11, 2026; and Jarvis Labs pricing, pause, resume, storage, filesystem, and Reserved IP rules were rechecked on September 14, 2026. HostFleet did not measure invoices, transition timing, availability, or performance.

A100 rental price per hour in 2026: 22 public rates checked

The cheapest selected A100 40 GB list rate is now Jarvis Labs at $0.89 per GPU-hour, or an estimated $640.80 for 720 active hours. That remains below Verda’s newly lower $1.251/hour row.

For A100 80 GB, Thunder Compute still has the lowest minimum launchable total in this comparison: an estimated $1.25/hour after required CPU is added to its $1.09 GPU base rate. Hyperstack and Massed Compute publish complete one-GPU VM totals of $1.35/hour. Jarvis Labs lists $1.49/hour, followed by RunPod’s selected Secure Cloud PCIe Pod at $1.59/hour.

Those numbers do not buy the same thing. This comparison spans 40 GB and 80 GB cards, PCIe and SXM4 systems, complete VMs, Pods, per-second containers, managed deployments, and Northflank GPU components that still require paid CPU and memory. Choose the memory class and operating surface before sorting by price.

The September 11 correction is about product scope, not a rate change. Paperspace’s $3.09/hour A100 40 GB and $3.18/hour A100 80 GB rows are Machine compute prices. The default 50 GB SSD is configured with the Machine but billed separately at $0.0074/hour, capped at $5/month. Linux auto-shutdown can stop compute after a configured interval, but the public documentation does not define what counts as inactivity. Treat it as a watchdog to test, not proven load-aware scale-to-zero.

Price anchors rechecked: September 14, 2026
Paperspace lifecycle and add-on rules rechecked: September 11, 2026
Jarvis Labs lifecycle and add-on rules rechecked: September 14, 2026

Currency: public USD list rates before tax
Monthly estimate: one listed product held billable for 720 hours
Evidence boundary: a public rate does not prove current stock, quota, regional access, approval, or equal performance

The short answer

RequirementLowest selected public rate720-hour planning figureImportant boundary
A100 40 GBJarvis Labs at $0.89/hr$640.80One-GPU on-demand row; capacity is released when paused
A100 80 GB, lowest launchable totalThunder Compute at $1.25/hr estimated$900.00$1.09 GPU base plus four required paid vCPUs
A100 80 GB, complete published VM totalHyperstack or Massed Compute at $1.35/hr$972.00Different regions, hardware variants, resources, and lifecycle rules
A100 80 GB, request-waking serviceKoyeb at $1.60/hr while active$1,152.00 if active for 720 hoursPublic-preview scale-to-zero; five-minute default idle period
Managed-cloud componentNorthflank at $1.42/hr for 40 GB or $1.76/hr for 80 GBGPU component onlyCPU, memory, disk, and egress are separate

All values in this summary come from the linked official rate sources and were rechecked September 14, 2026. The 720-hour figures are arithmetic, not vendor quotes.

A100 40 GB: five complete rates plus one component price

A100 40 GB is a separate capacity class. If the model, KV cache, batch, and runtime overhead do not fit with safe headroom, its lower rates are irrelevant.

Provider and productConfiguration or boundaryPublic rate720-hour estimateOfficial source and check date
Jarvis Labs on-demand instance1x A100 40 GB; public row lists 16 vCPU and 112 GB RAM$0.89/GPU-hr$640.80Jarvis Labs pricing, Sept. 14, 2026
Verda GPU instance1x A100 40 GB SXM4; 22 CPU and 120 GB RAM included; storage separate$1.251/hr$900.72Verda pricing, Sept. 14, 2026
Lambda Cloud VMSelected one-GPU A100 PCIe or SXM row; 40 GB$1.99/GPU-hr$1,432.80Lambda GPU instances, Sept. 14, 2026
Modal containerA100 40 GB allocated to a serverless container$0.000583/sec ($2.0988/hr)$1,511.14Modal pricing, Sept. 14, 2026
Paperspace Machine1x A100 40 GB; 12 vCPU and 90 GB RAM; default 50 GB SSD billed separately$3.09/hr compute$2,224.80 computePaperspace pricing and Machine Type Reference, Sept. 14, 2026

Northflank also publishes an A100 40 GB rate of $1.42/GPU-hour, checked September 14, 2026 on Northflank’s pricing page. That is $1,022.40 for 720 GPU-component hours, before the required CPU and memory compute plan. Because Northflank does not prescribe one minimum CPU/RAM shape for this card, inventing an all-in total would make the table look more precise than the source allows.

The estimates multiply the native rate by 720. Modal’s figure uses its unrounded per-second rate. They exclude separately billed storage, networking where charged, IPs, support, tax, additional replicas, and operational labor. Paperspace’s default disk is specifically outside its compute-only monthly figure.

A100 80 GB: 15 product rates plus one component price

This table ranks published product rates, not Northflank’s incomplete GPU component. Novita remains in the catalog-price comparison but is not currently launchable from the public inventory response. Thunder is the one ranked row whose minimum total requires transparent CPU arithmetic.

Provider and productConfiguration or access boundaryPublic rate used720-hour estimateOfficial source and check date
Thunder Compute1x A100 80 GB; minimum 8 vCPU, 64 GB RAM, and 100 GB disk$1.09 GPU base + 4 × $0.04 vCPU = $1.25/hr minimum VM$900.00Thunder pricing, pricing API, and spec API, Sept. 14, 2026
Hyperstack VM1x A100 80 GB PCIe; 28 CPU, 120 GB RAM, and local storage included; Canada$1.35/GPU-hr$972.00Hyperstack pricing, Sept. 14, 2026
Massed Compute VM1x A100 80 GB; 16 vCPU and 96 GB RAM included$1.35/hr$972.00Massed Compute pricing, Sept. 14, 2026
Jarvis Labs on-demand instance1x A100 80 GB; public row lists 16 vCPU and 112 GB RAM$1.49/GPU-hr$1,072.80Jarvis Labs pricing, Sept. 14, 2026
RunPod Secure Cloud PodSelected one-GPU A100 80 GB PCIe Pod; Secure Cloud SXM is also $1.59/hr$1.59/hr$1,144.80RunPod pricing, Sept. 14, 2026
Novita AI instance1x A100 80 GB SXM; 14 vCPU, 240 GB RAM, and 60 GB container-disk quota; API reports inventoryState=none, usableNode=false, and zero available GPUs$1.60/GPU-hr$1,152.00Novita marketplace API, Sept. 14, 2026
Koyeb GPU Service1x A100 80 GB; 15 vCPU, 180 GB RAM, and 320 GB disk included$1.60/hr$1,152.00Koyeb pricing, Sept. 14, 2026
Verda GPU instance1x A100 80 GB SXM4; 22 CPU and 120 GB RAM included; storage separate$1.736/hr$1,249.92Verda pricing, Sept. 14, 2026
Vultr Cloud GPU1x PCIe A100 80 GB catalog plan; 12 vCPU, 120 GB RAM, 1.40 TB local storage, and 10 TB bandwidth included; API locations array is empty$2.397/hr$1,725.84Vultr public plan API, Sept. 14, 2026
Modal containerA100 80 GB allocated to a serverless container$0.000694/sec ($2.4984/hr)$1,798.85Modal pricing, Sept. 14, 2026
CoreWeave InferenceA100 80 GB single-GPU inference rate; inference-platform customers only$2.70/GPU-hr$1,944.00CoreWeave pricing, Sept. 14, 2026
RunPod ServerlessA100 80 GB worker tier; not an exact-card reservation$2.72/hr$1,958.40RunPod pricing, Sept. 14, 2026
Paperspace Machine1x A100 80 GB; 12 vCPU and 90 GB RAM; default 50 GB SSD billed separately$3.18/hr compute$2,289.60 computePaperspace pricing and Machine Type Reference, Sept. 14, 2026
Baseten deploymentA100 80 GiB managed deployment$0.06667/min ($4.0002/hr)$2,880.14Baseten pricing, Sept. 14, 2026
Replicate private deploymentOne A100 80 GB managed deployment$0.001400/sec ($5.04/hr)$3,628.80Replicate pricing, Sept. 14, 2026

Northflank’s separate A100 80 GB component is $1.76/GPU-hour, or $1,267.20 for 720 GPU-component hours, from its official pricing checked September 14, 2026. The all-in running rate must add the chosen vCPU and memory plan; persistent disk and egress can add more. The raw GPU number sits between Koyeb and Verda, but the unknown complete total means it should not be ranked between them.

Together, the guide now contains six A100 40 GB cells and 16 A100 80 GB cells: 22 public price points. They are not 22 equivalent rentals.

The Vultr A100 plan is included because the official plan API returned it again on September 7 after it was absent during the September 4 review. Its current record has an empty locations array, so this is volatile catalog evidence, not proof that an account can deploy it in any region.

What the five added cells change

The September 8 refresh expanded this page from 17 to 22 price cells. That revision added Jarvis Labs at both memory tiers, Koyeb at 80 GB, and Northflank components at both tiers.

Jarvis Labs leads 40 GB, but pause is not free storage

Jarvis Labs publishes $0.89/GPU-hour for A100 40 GB and $1.49/GPU-hour for A100 80 GB, rechecked September 14, 2026 on its pricing page. The selected on-demand rows list one GPU, 16 vCPU, and 112 GB RAM. The page says on-demand instances bill per minute, but neither the pricing page nor the checked FAQ defines whether partial minutes are rounded, prorated, or aggregated. Do not turn a sub-minute run into an invoice claim.

Pause changes the meter and the capacity promise. The Jarvis Labs SDK documentation, checked September 14, says a successful pause preserves installed packages and files while stopping compute billing. The API accepts the request before the instance finishes moving through Pausing to Paused; the public sources do not establish the exact point inside that transition when compute billing stops.

Once paused, instance data costs $0.00014/GB-hour, according to the Jarvis Labs FAQ checked September 14. That is $0.007/hour for 50 GB and the vendor’s own $5.04 for 720 hours example. The pricing page rounds storage to $0.10/GB-month; the hourly meter produces $0.1008 per GB over 720 hours. Use the hourly rate for lifecycle estimates and preserve the displayed monthly headline as a rounded planning number.

Here is the intermittent-work boundary for one 50 GB instance:

Jarvis Labs A10024 active hours696 paused hoursDerived 30-day total720 active hours
A100 40 GB$21.36 compute$4.872 storage$26.23$640.80
A100 80 GB$35.76 compute$4.872 storage$40.63$1,072.80

Estimate assumptions: one instance, 24 whole active hours, 696 fully paused hours, 50 GB of instance storage, no overlap during state transitions, no shared filesystem, no Reserved IP, no tax, and unchanged September 14 rates. The calculations are active rate × 24 + $0.00014 × 50 × 696, rounded to cents only at the end. They are not measured bills, and they exclude any unresolved partial-minute or transition treatment.

Pause also releases the GPU. The FAQ says a later resume is not guaranteed, while the SDK says resume remains tied to the original region, can fail when that GPU is unavailable there, and may return a new machine_id. Automation must wait for Paused, store the ID returned by resume, and treat the retained disk as recoverable working state rather than reserved compute capacity.

Destroy has narrower cleanup semantics than its name suggests. The SDK says destroying an instance permanently deletes that instance and its instance storage. A shared filesystem is an independent provisioned-capacity resource at $0.00014/GB-hour, checked September 14, and survives instance termination until separately removed. A Reserved IP is VM-only and available only in IN1 and IN2, not EU1. For users outside the Indian-user billing category, it carries a separate, non-refundable $5 charge per 30-day cycle; destroying the VM returns the reservation to Idle, and billing continues until the IP is explicitly released.

Finally, paused storage is not a backup. Jarvis warns that a zero wallet can release all resources and permanently delete retained data. Export important artifacts before pausing, keep the wallet funded for anything intentionally retained, and delete instance storage, shared filesystems, and Reserved IPs explicitly when the job is finished.

Koyeb adds an included-resource service with a real idle tail

Koyeb lists one A100 80 GB Service at $1.60/hour, with 15 vCPU, 180 GB RAM, and 320 GB disk on the pricing card, rechecked September 14, 2026. The separate A100 SXM product is $2.15/hour and is not substituted into the lower selected row.

Koyeb’s scale-to-zero documentation, checked August 28, 2026, explicitly includes GPU Instances. A Service can set its minimum instance count to zero, with a default five-minute idle period. At the current $1.60/hour A100 rate, the nominal five-minute post-traffic tail is:

$1.60 × 5 / 60 = $0.1333

That is derived planning arithmetic, not a vendor-quoted minimum charge. It excludes active processing, wake and model-load time, additional instances, storage, and networking. Scale-to-zero is in public preview; held connections prevent idleness, HTTP/2 cannot wake a sleeping Service, and no public GPU wake-latency SLA was found.

For a sparse request-driven endpoint, the active hourly rate plus the five-minute tail can matter more than a 720-hour estimate. For a continuously warm endpoint, it behaves like the $1,152 monthly planning row.

Northflank adds useful market evidence, not a finished invoice

Northflank publishes $1.42/GPU-hour for A100 40 GB and $1.76/GPU-hour for A100 80 GB, rechecked September 14, 2026. Its managed-cloud documentation says GPU use is billed by the second once provisioned.

The trap is that these are GPU components. Northflank separately charges for the selected CPU and memory compute plan, and also lists persistent disk and egress charges. Its public documents do not prescribe a model-specific minimum CPU/RAM plan that would justify one comparable all-in figure.

Northflank’s autoscaling is also not request-waking scale-to-zero. The manual scaling documentation, checked August 31, 2026, says a service can be manually set to zero but is unavailable at zero. Its autoscaling documentation describes 15-second evaluations and a five-minute moving downscale window, but does not establish an autoscaling minimum of zero or an incoming-request wake path. Treating it as interchangeable with Koyeb would erase the central product difference.

Forty gigabytes versus 80 GB is the first decision

Jarvis’s $0.89 A100 40 GB rate is 36 cents below Thunder’s estimated $1.25 minimum 80 GB VM total. That gap is real, but it is useful only if the workload fits.

Start with model weights, precision, KV-cache size, batch size, context length, runtime workspace, and headroom for the serving stack. If the safe requirement exceeds 40 GB, remove every 40 GB row. HostFleet’s Llama 70B VRAM guide shows the memory arithmetic, including why context and concurrency can consume the apparent spare capacity.

Hardware variant is the next filter. PCIe and SXM4 systems have different bandwidth and topology boundaries. A single-GPU inference service may care less about multi-GPU fabric than training or tensor-parallel inference, but the exact variant still belongs in the deployment record. “A100” alone is not a reproducible configuration.

What the 720-hour estimates mean

A 30-day planning month has 720 hours. The estimates assume one named product stays billable for all 720 hours. They normalize continuously allocated capacity; they do not predict a bursty endpoint’s bill.

Vultr’s GPU billing guide uses 730 hours for its published monthly calculation while charging the actual hours in each calendar month. This guide applies 720 hours uniformly, so the Vultr row’s $1,725.84 planning figure differs from the API’s $1,750 catalog monthly price.

The calculations assume:

  • one GPU product remains billable continuously;
  • no overlapping rollout, failed replacement, or extra replica is charged;
  • native per-second or per-minute rates are multiplied before display rounding; and
  • the public list rate remains unchanged for the planning period.

They exclude storage beyond included allocations, network overages, public IPs, support, tax, commitments, regional premiums, retries, and engineering labor. Thunder’s $900 estimate includes the minimum required vCPU arithmetic. Northflank’s figures deliberately remain GPU-component floors. Paperspace’s displayed monthly values are compute-only because its default disk is a separate charge.

For intermittent work, billable allocation time replaces 720 in the formula. Startup, image pulls, model loading, retries, idle windows, minimum workers, downscale holds, and retained resources all change cost. HostFleet’s GPU cloud cost calculator exposes useful-hour and always-warm assumptions instead of burying them inside one monthly headline.

Paperspace: the default disk is not included, and auto-shutdown is not proven scale-to-zero

Paperspace’s September 11 documentation check corrects two easy assumptions about the A100 rows.

First, the A100 Machine type includes 12 vCPU, 90 GB RAM, and a default 50 GB SSD in its configuration, but the pricing page lists storage separately. The default disk costs $0.0074/hour with a $5 monthly cap, checked September 11, 2026. A public IP is another line item at $0.0045/hour with a $3 monthly cap. A dynamic public IP exists only while the Machine is on; a static IP persists and bills until it is deleted.

That changes the continuously-on planning totals:

Paperspace A100 Machine720-hour computeDefault 50 GB diskPublic IP if held for the monthMaximum listed bundle estimate
A100 40 GB$2,224.80$5.00 monthly capup to $3.00 monthly capup to $2,232.80
A100 80 GB$2,289.60$5.00 monthly capup to $3.00 monthly capup to $2,297.60

Estimate assumptions: 720 powered-on hours at the September 14 compute rates; one default 50 GB disk retained for the month; one public IP billed long enough to hit its monthly cap; no tax, snapshots, private networking, or other add-ons. The IP is optional, so the disk-adjusted figures without it are $2,229.80 and $2,294.60. These are arithmetic planning totals, not invoice measurements.

Second, Off is the only documented Machine state without hourly usage fees. Provisioning, Starting up, On/Ready, and Shutting Down should not be treated as free. Paperspace says a provider stop takes approximately one minute, but the public pricing documentation says only that Machines are billed per hour; it does not disclose whether partial hours are prorated, rounded, or metered more finely. Do not turn a 75-minute test into an invoice prediction by multiplying the hourly rate alone.

Linux auto-shutdown is useful, but its semantics are not documented well enough to call it request-driven scale-to-zero:

  • the selectable delay is one hour to one week;
  • a Linux Machine shuts down after the selected inactivity period even when users are connected;
  • the public guide does not define inactivity or state whether GPU use, CPU use, SSH traffic, an open SSH session, or inference requests reset the timer; and
  • enabling or disabling auto-shutdown on an existing Machine is console-only, not available through the Paperspace API or CLI.

For a development box, test the one-hour setting with a noncritical workload and an external stop watchdog. For a production inference server, use explicit API-driven lifecycle automation until a workload-specific test proves what the inactivity timer observes. Power-off ends Machine compute, not every charge: the disk and a static IP continue billing. Deactivation removes the Machine, its files, and its snapshots permanently and requires the Machine to be off first.

The off switch can reverse the ranking

The cheapest hourly row is not always the cheapest operational choice.

  • Jarvis Labs: wait for Paused before assuming compute has ended; instance storage then bills at $0.00014/GB-hour and capacity is released. Resume is region-locked and capacity-dependent. Destroy removes instance storage, not an independent shared filesystem or an unreleased Reserved IP.
  • Thunder Compute: its billing guide says compute bills per minute while the instance runs and deletion stops instance billing. Confirm retained-disk treatment.
  • Hyperstack: its states-and-billing guide says a stopped VM remains billable because hardware stays reserved; hibernation deallocates the flavor, while retained resources can still bill.
  • Massed Compute: its billing overview says the active VM total is debited per minute. Confirm the exact release action before automation.
  • RunPod Pods: compute allocation and persistent storage have separate lifecycle controls.
  • Novita: its GPU instance pricing guide says stopping ends compute billing; storage is separate.
  • Koyeb: an eligible Internet-facing GPU Service can scale to zero after its idle window, subject to preview and protocol limitations.
  • Verda: its instance lifecycle guide says shutdown does not stop compute billing; deletion is required, and retained storage can continue billing.
  • Vultr: its stopped-instance billing documentation says a stopped instance remains billable until it is destroyed.
  • Paperspace: only Off is documented as free of Machine usage fees. Power-off stops compute, but the default disk, a static IP, and other retained add-ons can continue billing; Linux auto-shutdown is not documented as load-aware.
  • Northflank: manual zero makes the service unavailable; documented autoscaling does not establish request wake-up from zero.
  • Modal and RunPod Serverless: workers can return to zero, but startup, execution, idle windows, and warm settings determine billable allocation.

A portable cleanup job cannot merely call “stop” everywhere. For each provider, record the exact action that releases the GPU and the fate of disks, checkpoints, images, IPs, and cached model data.

Choose the operating surface after memory

Once the memory floor is fixed, group products by the work they remove.

  1. Self-managed VM or Pod: Jarvis Labs, Thunder, Hyperstack, Massed Compute, RunPod Pods, Novita, Verda, Vultr, Lambda, and Paperspace leave the image, inference server, authentication, rollout, monitoring, and cleanup largely to the buyer.
  2. Request-driven or scale-to-zero container: Koyeb, Modal, and RunPod Serverless can align compute with active allocation, but their wake path, idle policy, and worker semantics differ.
  3. Managed deployment: Baseten, Replicate, and eligible CoreWeave products add a more opinionated serving surface. Higher rates can be rational when autoscaling, rollout controls, and operations replace internal work.
  4. Component-priced managed cloud: Northflank exposes a GPU component inside a configurable service. Build the complete CPU, memory, storage, and egress total before comparing it with fixed VM bundles.
  5. Catalog evidence: a public price is not current capacity. Account access, quota, region, and actual launch success remain separate checks.

The serverless GPU pricing matrix is the better companion for worker and managed-deployment tradeoffs. RunPod’s pricing guide explains why Pods and Serverless need different allocation and storage assumptions. If A100 is not the required generation, compare the H100 rental price guide after fixing the workload shape.

A practical A100 buying checklist

  1. Set the memory floor. Remove every 40 GB row if weights plus runtime headroom require 80 GB.
  2. Confirm the exact hardware. Record PCIe versus SXM4, topology, region, and the smallest deployable GPU count.
  3. Build the complete product total. Add required CPU, memory, storage, IP, and egress components.
  4. Prove capacity. Check account eligibility, quota, approval, region, and live inventory.
  5. Test the off switch. Verify whether pause, stop, hibernate, scale-to-zero, manual zero, or deletion ends compute billing.
  6. Run a bounded deployment test. Measure provisioning, model load, billed duration, failure recovery, and cleanup before production traffic.
  7. Set a spend guardrail. Alert on unexpected replicas and retained resources; do not rely on a low hourly rate to limit a broken rollout.

Verdict

Jarvis Labs remains the selected A100 40 GB public-rate leader at $0.89/GPU-hour, or an estimated $640.80 for 720 active hours, based on the official rate checked September 14, 2026. For intermittent work, 24 active hours plus 696 paused hours with 50 GB of instance storage is an estimated $26.23, but pause releases the GPU and does not stop storage billing. Shared filesystems and Reserved IPs require their own cleanup.

Thunder Compute remains the lowest A100 80 GB launchable total in this check at an estimated $1.25/hour. That is explicit component arithmetic: a $1.09 GPU base plus four paid vCPUs at $0.04 each. It is not a vendor-published all-in headline or proof of inventory.

Hyperstack and Massed Compute tie at $1.35/hour among selected published complete VM totals. Jarvis Labs is $1.49/hour, RunPod’s selected Secure Cloud PCIe Pod is $1.59/hour, and Koyeb is $1.60/hour with an included-resource, request-waking Service boundary.

Northflank’s $1.42 and $1.76 A100 numbers are GPU components, not comparable totals. They are useful public market data only after CPU and memory are added.

Paperspace’s $3.09 and $3.18 A100 rates are compute-only, not complete Machine totals. Add the separately billed default disk, then any IP or other retained resource. Auto-shutdown can cap a development session, but the undefined Linux inactivity signal means it should not be sold as production scale-to-zero.

The defensible buying order is memory, exact hardware, product shape, capacity, billing lifecycle, complete cost, and then hourly rate. The lowest number wins only after every earlier constraint survives.

Sources

  • HostFleet GPU pricing dataset — live-table baseline; its last full-dataset check covered all 21 providers and 148 displayed price cells on September 10, 2026
  • Jarvis Labs pricing — A100 40 GB and 80 GB on-demand rates, per-minute billing statement, resource rows, and rounded storage headline; rechecked September 14, 2026
  • Jarvis Labs FAQ — per-minute usage language, paused instance-storage rate, 50 GB example, capacity release, zero-wallet behavior, and data-loss warning; checked September 14, 2026
  • Jarvis Labs SDK — asynchronous pause, compute-billing stop statement, retained instance state, capacity-dependent region-locked resume, possible replacement machine ID, destroy semantics, and independent filesystems; checked September 14, 2026
  • Jarvis Labs CLI — machine-readable lifecycle fields, automatic post-run pause, and command boundaries; checked September 13, 2026
  • Jarvis Labs shared-filesystem documentation — independent lifecycle and $0.00014/GB-hour provisioned-capacity billing; checked September 14, 2026
  • Jarvis Labs Reserved IP documentation — VM-only scope, IN1/IN2 availability, Indian-user versus other-user billing categories, non-refundable 30-day charge, persistence after VM destroy, and explicit-release requirement; checked September 14, 2026
  • Verda pricing — exact A100 40 GB and 80 GB on-demand rates of $1.251/hr and $1.736/hr plus fixed CPU/RAM configurations; rechecked September 14, 2026
  • Lambda GPU instances — selected one-GPU A100 40 GB rate; rechecked September 14, 2026
  • Modal pricing — A100 40 GB and 80 GB per-second rates; rechecked September 14, 2026
  • Paperspace pricing — A100 compute rates, Off-state billing boundary, separately billed block storage, public-IP rates, and bandwidth; rechecked September 14, 2026
  • Paperspace Machine Type Reference — 12-vCPU, 90-GB-RAM, and default 50-GB-SSD configurations for A100 and A100-80G; rechecked September 14, 2026
  • Paperspace auto-shutdown guide — one-hour-to-one-week range, Linux connected-user behavior, and console-only control; checked September 11, 2026
  • Paperspace Machine states — billable-state boundary and approximate stop duration; checked September 11, 2026
  • Paperspace Machine features and public-IP guide — dynamic and static IP lifecycles; checked September 11, 2026
  • Paperspace Machine creation guide — separate Machine, disk, and feature price-summary boundary; checked September 11, 2026
  • Paperspace deactivation guide — off-state prerequisite and permanent deletion of the Machine, files, and snapshots; checked September 11, 2026
  • Thunder pricing, pricing API, and spec API — A100 base rate, vCPU price, and selectable minimum; rechecked September 14, 2026
  • Hyperstack pricing — selected one-GPU A100 80 GB PCIe VM rate; rechecked September 14, 2026
  • Massed Compute pricing — selected one-GPU A100 80 GB VM total; rechecked September 14, 2026
  • RunPod pricing — Secure Cloud PCIe/SXM and Serverless A100 rates; rechecked September 14, 2026
  • Novita marketplace API — A100 configuration and catalog rate, plus zero inventory, inventoryState=none, and usableNode=false; rechecked September 14, 2026
  • Koyeb pricing — A100 and A100 SXM rates and included resources; rechecked September 14, 2026
  • Koyeb scale-to-zero documentation — GPU eligibility, idle period, wake path, preview status, and protocol limits; checked August 28, 2026
  • Northflank pricing — A100 GPU-component, CPU, memory, storage, and egress pricing; rechecked September 14, 2026
  • Northflank managed GPU documentation — component and provisioning boundary; checked August 31, 2026
  • Northflank autoscaling and manual scaling — downscale window and manual-zero behavior; checked August 31, 2026
  • Vultr public plan API — restored A100 catalog record, rate, included resources, and empty locations array; rechecked September 14, 2026
  • Vultr GPU billing guide — 730-hour monthly calculation and actual calendar-month billing; rechecked September 14, 2026
  • Vultr stopped-instance billing — stopped-instance resource reservation and destroy-to-end-billing rule; rechecked September 14, 2026
  • CoreWeave pricing — A100 single-GPU inference rate and eligibility; rechecked September 14, 2026
  • Baseten pricing — A100 80 GiB managed-deployment rate; rechecked September 14, 2026
  • Replicate pricing — A100 80 GB private-deployment rate; rechecked September 14, 2026
  • Local evidence: /opt/hostbot-v2/src/data/gpu-pricing.json, /opt/hostbot/data/ai-hosting/notes/2026-09-10-gpu-pricing-full-verification.md, /opt/hostbot/data/ai-hosting/notes/2026-09-11-paperspace-linux-auto-shutdown-cost-boundary.md, /opt/hostbot/data/ai-hosting/notes/2026-09-13-jarvislabs-pause-destroy-billing-boundary.md, /opt/hostbot/data/ai-hosting/notes/2026-08-24-jarvis-labs-gpu-pricing.md, /opt/hostbot/data/ai-hosting/notes/2026-08-28-koyeb-gpu-scale-to-zero-limits.md, and /opt/hostbot/data/ai-hosting/notes/2026-08-31-northflank-gpu-autoscaling-billing-boundary.md

Need a self-managed A100 endpoint? This is a labeled affiliate link; source citations above remain direct. RunPod signup (+$5 credit on your first $10, affiliate) supports HostFleet’s testing budget at no extra cost to you. Re-check the exact card, region, rate, storage, and shutdown behavior before purchase.