Skip to content
GPU Finder

RTX 5090 cloud GPU pricing and rental comparison

As of , the cheapest single RTX 5090 confirmed in stock is $0.60 per GPU-hour on Lium. The lowest listed price is $0.40 per GPU-hour on Vast, but that provider does not expose a current stock signal. GPU Finder compares 6 providers for this size and refreshes availability hourly. Methodology.

Compare RTX 5090 cloud rentals for local LLM inference, image generation, and short fine-tuning runs across 6 providers.

Providers:6
Whole-node range:$0.40 — $13.87/hr
Spot price:from $0.27/hr
Instances:241
Available:1 providers
Free egress:2 providers
Updated:

Structured offer basis: 73 current 1× RTX 5090 listings range from $0.40 to $1.20 USD per physical GPU-hour. Whole-node prices for other GPU counts are not mixed into this range.

Cheapest single RTX 5090 rental

verified

Cheapest RTX 5090 confirmed in stock right now

Limited
$0.60/hr1× listing
on Lium · Ukraine
Visit Lium
Is now a good time to rent an RTX 5090 cloud GPU?
■ mid-range vs the last 6 months — The 1× RTX 5090 on-demand floor is $0.32/hr ($0.32/GPU-hr) — 6-month range $0.20–$0.65/hr for 1× nodes. In short, prices are mid-range versus the last 6 months.
What is the cheapest RTX 5090 cloud GPU available right now?
Cheapest RTX 5090 confirmed in stock right now: $0.60/GPU-hr on Lium (1× listing, $0.60/hr).
Which cloud providers most reliably have the RTX 5090 in stock?
Most reliably in stock over the last 30 days: Vast. Frequently waitlisted: Theta EdgeCloud, Runpod.

Prices are the cheapest listed on-demand rates; “in stock” means the provider reported capacity on our last hourly check. How we compute this →

RTX 5090 rental prices by provider

  • 1x RTX 5090 marketplace
    Stale · 7h ago0% · 30d
    $0.40/hr
    Spot $0.38−6%
    1× RTX 509032 vCPU79 GBFree egressCzechia, CZ

Get notified when RTX 5090 drops in price or reliability changes

One email per change, max once a day. We send a confirmation link first; one-click unsubscribe in every email.

RTX 5090 (1× GPU) availability — last 7 days

No provider held 1× RTX 5090 ≥80% available across all 7 days. Polled hourly · 5 providers tracked · 7-day window.

≥80%50–80%<50%No data
View:
Provider
29
30
1
2
3
4
5
Week
Shadeform51%
Lium20%
Vast0%
Runpod0%
Theta EdgeCloud0%

30-day reliability depth

GPU Finder does not treat availability as a static yes/no flag. For RTX 5090, current stock badges are paired with a 30-day reliability score based on our hourly stock checks — the share of tracked time each listing was reported available. How we compute this →

223
listings with a 30-day score
93%
of listings on this page have a score
30d
longest stock history on this page

Strongest reliability signals here: Vast (89%, 8d covered), Lium (71%, 30d covered), Shadeform (30%, 30d covered).

Caveat: scores stay hidden until a listing has at least 48 tracked hours, and some provider APIs expose coarse capacity levels instead of exact stock counts. Use the score as historical depth beside current availability, not as a guarantee that a GPU will still be allocatable when you click through.

Loading chart...

Spot pricing decision guide

Treat spot as a risk-adjusted capacity decision, not just a cheaper number.

Sparse spot history
On-demand floor
$0.34

per GPU-hour

Interruptible floor
$0.21

37% below on-demand floor

7-day trend
flat

daily median for the same offers with complete history

Volatility
stable

median hidden until ≥3 rows across ≥2 providers

Cheapest trusted interruptible capacity is on Vast; 0 providers currently report spot or interruptible stock.

Vast has kept spot or interruptible capacity in stock most reliably over the last 30 days (89%). See the availability section for the broader stock history.

There is a valid spot or interruptible row, but not enough trusted history to overstate the signal.

Good fit: checkpointed batch jobs, flexible training, and experiments that can restart. Avoid for production serving, deadline-bound runs, or jobs that cannot tolerate eviction.

Source caveat: Runpod exposes explicit spot fields; Vast is labeled interruptible/bid-floor. Lambda is treated as on-demand-only unless a verified spot field is added.

About the RTX 5090

The NVIDIA RTX 5090 is the fastest consumer GPU commonly rentable in the cloud. Its 32 GB of GDDR7 gives local-LLM users more KV-cache headroom than a 24 GB RTX 4090 while keeping the hourly rate far below H100-class datacenter cards. It is best treated as a single-GPU inference and prototyping card: excellent for quantized 7B-34B models, less suitable for production clusters that need ECC, NVLink, or managed enterprise SLAs.

Key Specifications

ArchitectureBlackwell (GB202)
GPU Memory32 GB GDDR7
Memory Bandwidth1.79 TB/s
FP16 Tensor CoreBlackwell Tensor Cores with FP4 support
TDP575W
InterconnectPCIe Gen5 x16 (no NVLink)
Release Year2025

Cloud Pricing Context

RTX 5090 cloud pricing is still concentrated among neocloud and marketplace providers rather than hyperscalers. The value proposition is consumer-GPU economics: rent the 32 GB card by the hour for local LLMs, benchmarks, image generation, and LoRA experiments instead of buying scarce desktop hardware up front. Compare the live table below against RTX 4090 rates when 24 GB is enough, and against L40S or H100 pricing when you need ECC, enterprise reliability, or larger VRAM.

Best For

  • Single-GPU local-LLM inference for 7B-34B quantized models
  • Longer context windows or larger batches than fit comfortably on a 24 GB RTX 4090
  • Short benchmark, evaluation, and LoRA/QLoRA experiments before buying hardware
  • Image and video generation workloads that benefit from Blackwell consumer throughput

RTX 5090 cloud rental at a glance

Fast checks for consumer-GPU buyers and local-LLM users deciding whether to rent a 32GB card.

Lowest per-GPU-hour
$0.34

Normalized from live RTX 5090 on-demand rows by GPU count.

VRAM class
32GB

8GB more headroom than RTX 4090 for KV cache and larger batches.

Best fit
Local LLMs

Single-GPU inference, evaluation, image generation, and light fine-tuning.

Before renting an RTX 5090

Use the live table as a shortlist, then confirm the actual checkout configuration with the provider.

Jump to live RTX 5090 prices
  • Treat $0.34/GPU-hr as a normalized per-GPU floor. A cheaper row can be a whole multi-GPU listing, so the billed configuration may be the full instance rather than a single RTX 5090.
  • Do not read the listed minimum as a guarantee that single-GPU RTX 5090 capacity is available. Check the table's stock label and the provider's current checkout state.
  • Verify region, CPU/RAM, local storage, storage and egress charges, and the software image or driver/CUDA support before starting the rental.
  • Match the 32GB VRAM limit to your quantization level, context length, and batch size; model fit changes with those choices.

At $0.34/GPU-hr, RTX 5090 rentals are the cloud version of trying a high-end desktop card before buying one: use them for bursty consumer-GPU jobs, then step up to L40S, H100, or H200 when you need ECC memory, NVLink-class scaling, or production reliability instead of the lowest single-card hourly rate.

RTX 5090 cloud GPU FAQ

How much does it cost to rent an RTX 5090 in the cloud?

GPU Finder currently tracks RTX 5090 cloud rental pricing from $0.34/GPU-hr across 6 providers. The table below shows live provider, region, availability, and on-demand or spot rates for 32GB consumer-GPU hosts; normalized per-GPU floors may come from multi-GPU listings that bill as a whole instance.

Is an RTX 5090 good for local LLMs?

Yes, when the model and workload fit in 32GB of VRAM after accounting for quantization, context length, and batch size. RTX 5090 instances are strongest for single-GPU local-LLM inference, image generation, evaluation, and short LoRA or QLoRA experiments that do not need datacenter interconnect.

Should I rent an RTX 5090 or an RTX 4090?

Choose RTX 5090 when the extra 8GB of VRAM or Blackwell throughput changes the job: longer context, larger batches, or a quantized model that needs more than 24GB without assuming every 32B-class setup fits. Choose RTX 4090 when 24GB is enough and the lower hourly rate matters more.