12 entries · Updated 2026-10-05

Rented GPUs

Where to rent the card you do not own: hourly, per-second and reserved GPU providers, with the billing unit, the regions and the price reference read from their own pages.

Lists

12 providers listed

per-second

RunPod

Secure Cloud and a community pool of hosts, both billed by the second, plus serverless endpoints and multi-node clusters.

Lists
B300 · B200 · H200 · H100 · A100 · L40S · A40 · RTX 6000
Billing
per second, no minimum
Minimum
1 GPU
Regions
Secure Cloud in EU and US datacenters, plus community hosts

Read on 2026-10-05: per-second prices span about $0.58 to $7.89 per GPU-hour across the catalogue. Community hosts are cheaper and appear and disappear.

runpod.io/pricing
per-second

Vast.ai

A marketplace rather than a datacenter: hosts rent out their own cards, rates follow supply and demand and are refreshed hourly.

Lists
B200 · H200 · H100 · A100 · RTX PRO 5000 · RTX 5090 · RTX 4090
Billing
per second; on-demand, interruptible or reserved
Minimum
1 GPU
Regions
40+ datacenters, host location varies

Medians read on 2026-10-05 run from about $0.45/GPU-hour on consumer cards to $7.81 on the newest datacenter ones. Your host is an unknown third party: the machine holds your weights and reads your prompts.

vast.ai/pricing
per-second

Modal

Serverless GPUs: nothing runs between requests, the container starts with the call and the meter stops when it exits.

Lists
B300 · B200 · H200 · H100 · A100 · L40S · L4 · A10G · T4
Billing
per second of execution, plus a monthly free credit
Minimum
one container at a time
Regions
their cloud, several regions

Read on 2026-10-05: B300 at $0.001972 per second, which is about $7.10 per GPU-hour. Cheapest per-second rate of the set, and the one that punishes a job that idles.

modal.com/pricing
per-minute

Hyperstack

H100, A100 and L40 virtual machines billed by the minute, with reserved private clusters for anything larger.

Lists
B200 · GB200 · B300 · H100 · A100 · L40 · A6000
Billing
per minute; reserved clusters by contract
Minimum
1 GPU
Regions
United Kingdom and continental Europe

Read on 2026-10-05: H100 SXM $3.20/GPU-hour, H100 NVLink $2.60, A100 SXM 80 GB $1.60, A100 NVLink $1.40; reserved H100 from $1.75. B300 was announced as coming, not listed.

hyperstack.cloud/pricing
per-second

OVHcloud

European public cloud with GPU instances, Bare Metal GPU and job-based AI Training, on a per-second or per-minute grid depending on the product.

Lists
H200 · H100 · A100 · L40S · L4 · V100
Billing
hourly list price, billed per second or per minute
Minimum
1 GPU
Regions
France (Roubaix, Gravelines) and other EU, plus Canada, Singapore, Sydney, Mumbai

The prices page is a comparison tool: pick the instance there. The same account can hold your VPS and its mail, so one invoice covers both.

ovhcloud.com/en/public-cloud/prices/
per-hour

Scaleway

French cloud renting GPU instances from a single L4 up to eight H100 in one instance, with the hourly price and a monthly estimate on the same line.

Lists
H100 · L40S · L4
Billing
per hour, monthly estimate shown
Minimum
1 GPU
Regions
France (Paris)

Read on 2026-10-05: L4 from €0.79/hour, L40S €1.58, H100 from €3.15, up to 8 GPUs per instance.

scaleway.com/en/pricing/gpu/
per-hour

Verda

Finnish AI cloud, formerly DataCrunch: single GPUs up to 8-GPU NVLink instances, spot and reserved rates, and instant clusters on InfiniBand.

Lists
B200 · H200 · H100 · A100 · L40S · A6000 · V100
Billing
per hour, on-demand, spot or reserved
Minimum
1 GPU
Regions
Finland (EU)

Read on 2026-10-05: on-demand from about $1.15/GPU-hour up to $10.32, spot below on-demand, and clusters up to 144 GPUs over InfiniBand. The site now brands itself Verda.

datacrunch.io/pricing
per-hour

Nebius

European cloud listing H100, H200, B200 and GB200 with managed Slurm and Kubernetes, and smaller prices on committed capacity.

Lists
GB200 · B200 · H200 · H100 · L40S
Billing
per hour on-demand; commitments advertised up to 35% lower
Minimum
1 GPU
Regions
EU

The prices page shows the GPU set and the discount structure; the per-hour figures sit behind the configuration picker.

nebius.com/prices
per-hour

Together AI

Rents GPU clusters by the hour alongside its hosted inference, including GB200 and GB300 NVL72 racks.

Lists
GB300 · GB200 · B200 · H200 · H100
Billing
per GPU-hour on-demand, reserved by contract
Minimum
1 GPU (clusters by contract)
Regions
United States

Two separate bills on one page: GPU time per GPU-hour, and tokens for the hosted models. Read on 2026-10-05 the hardware table listed GB200, GB300 and H-series racks.

together.ai/pricing
per-hour

Lambda

Instances, 1-Click Clusters and Superclusters, with the plainest hourly list prices of the set and reserved capacity on request.

Lists
B200 · H100 · A100 · A6000
Billing
per hour; reserved by contract
Minimum
1 instance
Regions
United States

Read on 2026-10-05: B200 SXM6 180 GB $6.69/GPU-hour, H100 SXM 80 GB $3.99, A100 SXM 80 GB $2.79, A100 SXM 40 GB $1.99.

lambda.ai/pricing
reserved

DigitalOcean

GPU Droplets next to the ordinary cloud furniture: H100, H200, B300 and MI300X, on-demand or on a twelve-month reserved plan.

Lists
B300 · H200 · H100 · MI300X · L40S
Billing
per second on-demand; 12-month reserved plans
Minimum
1 GPU
Regions
United States, Canada, Europe

Reserved prices read on 2026-10-05: B300 $7.94/GPU-hour, H200 $3.40, H100 $3.26. Spot droplets are interruptible with roughly two hours' notice.

digitalocean.com/pricing/gpu-droplets
reserved

CoreWeave

Large NVIDIA clouds with on-demand, spot and reserved capacity, GB200 NVL72 racks included, and managed Kubernetes that expects to be the whole estate.

Lists
GB200 · B200 · H200 · H100 · A100 · L40S · A6000 · A40
Billing
per hour on-demand and spot, reserved by commitment
Minimum
1 GPU (racks by contract)
Regions
see their site

Built for long commitments: the interesting prices are the reserved ones, and the page leads with the storage, networking and Kubernetes lines that come with them.

coreweave.com/pricing

What renting is for

A local rig answers most of the questions this site is about, and none of the ones that need more memory than you can buy. A 1M-token context on a 30B-class mixture-of-experts is about 90 GB of cache before the weights; a frontier MoE checkpoint is several hundred gigabytes. Both are rented problems, not upgrade problems.

Three things change when the card is someone else’s:

  • The data leaves the machine. Weights and prompts travel over the network and land on a machine you do not control. On a marketplace host you also do not know who that is. That is a decision, not a detail.
  • The meter runs on time, not on work. A per-second provider bills the container’s life, so a job that idles costs the same as a job that computes. A reserved contract bills whether you are there or not.
  • Storage and egress are separate lines. A 1 TB checkpoint that stays resident on their NVMe is billed monthly; pulling it out again is billed by the gigabyte. The 12 cards below list compute only.

How to read this page

Each card is a provider, and the two filter rows do the work: the first by billing unit (how fine the meter is), the second by which GPU families that provider actually lists. Prices are the providers’ own, read on the date inside each note — they move, and no aggregator figure is kept here. “Reserved” means the published headline price is a commitment, not a card you can have in an hour.