Skip to content
GPU Finder

L40S Cloud GPU Pricing

Compare L40S prices across 15 cloud providers. On-demand from $0.40/hr.

Providers:15
Price range:$0.40 — $43.70/hr
Spot price:from $0.20/hr
Instances:190
Available:8 providers
Free egress:6 providers
Updated:

L40S market snapshot

verified

Cheapest L40S confirmed in stock right now: $0.40/hr on Vast.

near its 6-month lowat $1.60/hr the L40S floor price is near its 6-month low — 6-month range $1.60–$3.40/hr.

Bottom line: a relatively good time to rent — the floor price is near its 6-month low.

Most reliably in stock over the last 30 days: PrimeIntellect, Shadeform. Frequently waitlisted: Cudo, Gcore.

30-day reliability depth

GPU Finder does not treat availability as a static yes/no flag. For L40S, current stock badges are paired with a 30-day reliability score computed from AvailabilityRecord history, measuring how much covered time each tracked row was reported available.

63
scored rows with enough history
33%
of visible rows have a 30-day score
30d
deepest covered window on this page

Strongest reliability signals here: PrimeIntellect (100%, 30d covered), Shadeform (92%, 30d covered), Vast (69%, 12d covered).

Caveat: scores stay hidden until a row has at least 48covered hours, and some provider APIs expose coarse capacity levels instead of exact stock counts. Use the score as historical depth beside current availability, not as a guarantee that a GPU will still be allocatable when you click through.

About the L40S

The NVIDIA L40S occupies the sweet spot between consumer GPUs and full datacenter accelerators. Based on Ada Lovelace, it provides 48 GB of GDDR6 with ECC and FP8 support in a standard PCIe, 350W form factor. It is purpose-built for inference-heavy deployments and workloads that need more memory than an RTX 4090 but do not justify H100-class pricing.

Key Specifications

ArchitectureAda Lovelace (AD102)
GPU Memory48 GB GDDR6 with ECC
Memory Bandwidth864 GB/s
FP16 Tensor Core362 TFLOPS
TDP350W
InterconnectPCIe Gen4 x16
Release Year2023

Cloud Pricing Context

L40S on-demand pricing starts as low as $0.32/hr on marketplace providers, with mainstream options around $1.57/hr. Its PCIe form factor means more providers can offer it without specialized NVLink infrastructure, keeping prices competitive.

Best For

  • Production inference serving for models in the 7B-13B range
  • Video and image generation pipelines (Stable Diffusion, Flux)
  • Mixed AI and visualization workloads using RT cores
  • Teams that need ECC memory at mid-tier pricing

Price History & Comparison

Loading chart...

Spot pricing decision guide

Treat spot as a risk-adjusted capacity decision, not just a cheaper number.

Cheap but risky
On-demand floor
$0.38

per GPU-hour

Spot / interruptible floor
$0.20

47% below on-demand floor

30-day trend
rising

fake/equal spot rows suppressed

Volatility
volatile

median $1.22/GPU-hr

Cheapest trusted spot / interruptible capacity is on Vast; 2 providers currently report spot-adjacent stock (113 units/signals visible).

Vast has the strongest 30-day spot-adjacent reliability signal here (69%). See the 30-day reliability depth module for the broader availability history.

The discount is visible, but availability, reliability, or volatility argues for more caution.

Good fit: checkpointed batch jobs, flexible training, and experiments that can restart. Avoid for production serving, deadline-bound runs, or jobs that cannot tolerate eviction.

Source caveat: Runpod exposes explicit spot fields; Vast is labeled interruptible/bid-floor. Lambda is treated as on-demand-only unless a verified spot field is added.

L40S (1× GPU) availability — last 7 days

No provider held 1× L40S ≥80% available across all 7 days. Polled hourly · 8 providers tracked · 7-day window.

≥80%50–80%<50%No data
View:
Provider
15
16
17
18
19
20
21
Week
AceCloud100%
PrimeIntellect88%
Shadeform58%
Nebius46%
Runpod33%
Vultr33%
Vast5%
Cudo0%

4× L40S server pricing and bundle math

Live-derived multi-GPU estimates for 4× L40S (192 GB aggregate VRAM).

Lowest per-GPU-hour
$0.38

Normalized from live on-demand rows by GPU count.

4× bundle per hour
$1.87

Cheapest live 4-GPU instance.

Estimated monthly
$1345.07

Approx. 24 × 30 at the bundle hourly rate.

GPU Finder currently tracks a live 4-GPU L40S instance from $1.87/hr ($1345.07/mo at 24×30). Use the table below to compare exact configurations and availability.

Networking, CPU, host memory, storage, and egress vary by provider and are not folded into this estimate. Prices are for on-demand cloud instances unless the provider table shows a spot or interruptible rate.

4× L40S server pricing FAQ

How much does a 4× L40S server cost per hour?

Based on current live rates, the cheapest 4× L40S bundle starts around $1.87/hr, which is roughly $1345.07/mo at 24×30 utilization. The query "4x nvidia l40s server price 2026" is the most common form of this search. L40S is a PCIe GPU, so 4× setups are usually workstation or small servers rather than NVLink clusters.

What is the monthly price for a 4× L40S server?

A 4× L40S server at $1.87/hr is about $1345.07/mo at 24×30. Add egress, storage, and CPU/RAM costs per provider.

How does 4× L40S compare to 4× RTX A6000?

Both 4× L40S and 4× RTX A6000 expose 48 GB of VRAM per GPU and use PCIe. L40S adds a Transformer Engine and FP8 support for inference, while RTX A6000 is a workstation card that can be cheaper for rendering and development workloads.

Where can I rent a single L40S GPU?

Use the live L40S provider table on this page. L40S is widely available from both neoclouds and mainstream cloud providers because its PCIe form factor lowers host complexity.

Is a 4× L40S server good for LLM inference?

Yes, for 7B–13B parameter models and image/video generation workloads that fit in 48 GB of GDDR6. L40S lacks NVLink, so large distributed training jobs are not its best fit, but it is one of the most cost-effective inference options with ECC memory.

Pricing by Provider

ProviderInstance TypeGPUsOn-DemandSpot / interruptibleVisit
Vast1x L40S marketplace1
$0.40
Limited
N/AVisit
Vast1x L40S marketplace1
$0.48
Limited
$0.47
Visit
Vast1x L40S marketplace1
$0.52
Limited
N/AVisit
Vast1x L40S marketplace1
$0.53
Limited
$0.20
Visit
Vast1x L40S marketplace1
$0.54
Limited
$0.33
Visit
Vast1x L40S marketplace1
$0.54
Limited
$0.47
Visit
Vast1x L40S marketplace1
$0.59
Limited
$0.43
Visit
Vast1x L40S marketplace1
$0.62
Limited
N/AVisit
Vast1x L40S marketplace1
$0.74
Limited
$0.47
Visit
Lium2x L40S marketplace2
$0.76
Available
N/AVisit
Runpod1x L40S Community Cloud1
$0.79
Limited
N/AVisit
Vast2x L40S marketplace2
$0.80
Limited
N/AVisit
Vast1x L40S marketplace1
$0.80
Limited
$0.77
Visit
Vast1x L40S marketplace1
$0.80
Limited
N/AVisit
Vast2x L40S marketplace2
$0.80
Limited
$0.67
Visit
SeewebECS1GPU71
$0.85
N/AVisit
Shadeformmassedcompute_L40S1
$0.88
Available
N/AVisit
Cudoepyc-genoa-l40s-graphics_1x2v4gb1
$0.90
Limited
N/AVisit
PrimeIntellectdatacrunch__1xL40S_48GB__20__601
$0.91
Available
N/AVisit
Packetpacket-l40s-us-east-va-dedicated-1gpu1
$0.92
N/AVisit

Get notified when L40S drops in price or reliability changes

One email per change, max once a day. We send a confirmation link first; one-click unsubscribe in every email.