One hour
$0.34
1 GPU × 1h · compute only
Ada cloud GPU
NVIDIA RTX 4090 provides 24 GB of GDDR6X and is a cost-focused choice for ComfyUI, Stable Diffusion, visual workloads and models that fit on one device. GPURento lists a $0.34 per GPU-hour rate with self-service creation in Paris and Frankfurt.
Reviewed by the GPURento infrastructure team · September 1, 2026
Cost scenarios
One hour
$0.34
1 GPU × 1h · compute only
One day
$8.16
1 GPU × 24h · compute only
Seven days
$57.12
1 GPU × 168h · compute only
Average month · 730h
$248.20
1 GPU × 730h · compute only
| Memory | 24 GB GDDR6X | Consumer GeForce configuration. |
|---|---|---|
| Architecture | Ada Lovelace | Strong tensor and graphics performance. |
| On-demand rate | $0.34 / GPU-hour | Compute only before persistent storage. |
| Available regions | Paris · Frankfurt | Selectable from the funded dashboard. |
Decision method
The content below separates published specifications, GPURento catalog references and decisions that still require a workload benchmark.
A 4090 is attractive only when the complete workload fits. Count model weights, activations, VAE, ControlNet, LoRA modules, KV cache and runtime overhead. CPU or disk offload can make a workflow run, but may erase the speed advantage.
For image and video work, compare seconds and dollars per accepted output at the same resolution, step count and quality settings. Hourly price without a repeatable workflow does not tell you which GPU is economical.
RTX 4090 is a GeForce product. Workloads that require ECC memory, data-center support characteristics or high-speed multi-GPU topology are better evaluated on A100, H100, H200 or L40S.
Catalog shortlist
24 GB GDDR6X · $0.34/hr
Image, video and mid-size inference
32 GB GDDR7 · $0.69/hr
Fast single-GPU inference and media
48 GB GDDR6 ECC · $0.79/hr
Production inference, fine-tuning and video
Questions
The catalog starts at $0.34 per RTX 4090 GPU-hour in Paris or Frankfurt. Storage and optional services are accounted for separately.
Yes, many ComfyUI image workflows fit within 24 GB. Large video models, high resolutions or multiple auxiliary models can exceed that limit, so test the exact graph.
RTX 5090 adds 32 GB of memory and a newer architecture at a higher reference rate. Choose it when the extra 8 GB removes offload or enables a larger workflow.
Last reviewed September 1, 2026. GPURento rates are current catalog references; provisioning state remains visible in the workspace, and manufacturer specifications do not substitute for workload benchmarks.
Continue the research
Deployment plan
Confirm that 24 GB is enough, fund the workspace and choose the exact container and region configuration.