One hour
$0.79
1 GPU × 1h · compute only
Ada data-center GPU rental
NVIDIA L40S combines 48 GB of GDDR6 ECC memory with Ada Tensor, RT and media engines. It fits production inference, fine-tuning, rendering and image or video pipelines that need more memory and data-center features than a GeForce GPU. GPURento lists L40S at $0.79 per GPU-hour in Paris and Frankfurt.
Reviewed by the GPURento infrastructure team · September 1, 2026
Cost scenarios
One hour
$0.79
1 GPU × 1h · compute only
One day
$18.96
1 GPU × 24h · compute only
Seven days
$132.72
1 GPU × 168h · compute only
Average month · 730h
$576.70
1 GPU × 730h · compute only
| Memory | 48 GB GDDR6 ECC | Official NVIDIA L40S configuration. |
|---|---|---|
| Architecture | Ada Lovelace | Tensor, RT and media acceleration in a data-center form factor. |
| On-demand rate | $0.79 / GPU-hour | Compute only before persistent storage. |
| Available regions | Paris · Frankfurt | Selectable through the funded dashboard. |
Decision method
The content below separates published specifications, GPURento catalog references and decisions that still require a workload benchmark.
The value of L40S is not one peak benchmark. It combines AI compute, 48 GB of ECC memory, ray-tracing hardware and three encode/decode engines. That makes it useful when one environment serves models and produces visual or video outputs.
NVIDIA lists L40S as PCIe Gen4 and does not list NVLink support. For tightly coupled multi-GPU training, compare H100, H200 or a cluster topology instead of assuming multiple L40S devices behave like one shared-memory accelerator.
RTX 5090 can offer attractive single-GPU throughput with 32 GB. L40S adds 48 GB, ECC and data-center positioning. Choose from the memory ceiling, reliability requirements and end-to-end output cost rather than generation alone.
Catalog shortlist
48 GB GDDR6 ECC · $0.79/hr
Production inference, fine-tuning and video
32 GB GDDR7 · $0.69/hr
Fast single-GPU inference and media
80 GB HBM3 · $2.69/hr
Intensive training and FP8 inference
Questions
GPURento lists L40S at $0.79 per GPU-hour. One continuous 24-hour day is $18.96 before storage, and 730 hours is $576.70.
The L40S has 48 GB of GDDR6 memory with ECC and an official memory bandwidth of 864 GB/s.
Choose L40S when 48 GB, ECC or data-center characteristics matter. RTX 5090 can be more attractive when the workload fits 32 GB and GeForce-class hardware is acceptable.
NVIDIA does not list NVLink support for L40S. Multi-GPU workloads should account for PCIe and host networking rather than assume an NVLink topology.
Last reviewed September 1, 2026. GPURento rates are current catalog references; provisioning state remains visible in the workspace, and manufacturer specifications do not substitute for workload benchmarks.
Continue the research
Deployment plan
Create a funded workspace and select the L40S image, region, runtime and persistent storage you need.