---
title: "Rent RTX 5090 — $0.439/hr Cloud GPU Pricing (2026) | PowerGPU"
description: "Rent the RTX 5090 (32 GB GDDR7) from $0.219/hr interruptible or $0.439/hr on-demand — fixed, ≥30% below market. 922 GPUs in 22 regions, per-second billing."
url: https://powergpu.ai/gpu/rtx-5090
last_modified: 2026-09-14T11:04:20+00:00
prices_as_of: 2026-09-14
site: PowerGPU (powergpu.ai)
---

Consumer flagship · Blackwell · launched 2025 · prices checked 2026-09-14

# Rent NVIDIA RTX 5090 — 32 GB, $0.439/hr on-demand

- VRAM 32 GB GDDR7
- FP16 tensor 419 TFLOPS
- PowerScore 295 RTX 3090 = 100
- Configs 1–8× PCIe 5.0
- Online now 922 22 regions

[On-demand (guaranteed) $0.439 /GPU-hr ≈ $320/mo · market ~~$0.63~~ (−30%)](https://cloud.powergpu.ai/?gpu=rtx-5090) [Interruptible $0.219 /GPU-hr flat −50% · pausable, disk kept](https://cloud.powergpu.ai/?gpu=rtx-5090&type=spot) [Reserved 3 mo $0.285 /GPU-hr −35% · capacity held for you](https://powergpu.ai/products/reserved)

Billed per second, price locked at deploy. Storage $0.08/GB/mo · bandwidth $0.01/GB — [the whole fee schedule](https://powergpu.ai/pricing). Next weekly market re-check: 2026-09-21.

Consumer flagship · Blackwell architecture

The RTX 5090 is the Blackwell consumer flagship: 32 GB GDDR7 at 1.79 TB/s, 21,760 CUDA cores and native FP4 in fifth-generation tensor cores. It is the best dollars-per-token card on the sheet for 7B–32B models and the entry point for video diffusion; supply is the deepest in the catalogue.

The community favourite: consumer pricing with serious tensor throughput. Ideal for diffusion models, quantized LLMs and fine-tuning runs that fit in 32 GB. Supply is deep, so interruptible capacity is almost always available at half price.

## NVIDIA RTX 5090 specs: VRAM, TFLOPS, bandwidth

- **GPU model**: NVIDIA RTX 5090 · **Architecture**: Blackwell (2025)
- **VRAM**: 32 GB GDDR7 · **Memory bandwidth**: 1,792 GB/s
- **FP16 tensor perf.**: 419 TFLOPS · **FP32 perf.**: 104.8 TFLOPS
- **CUDA cores**: 21,760 · **TDP**: 575 W
- **PowerScore ((RTX 3090 = 100))**: 295 · **PCIe generation**: Gen 5.0
- **Multi-GPU**: 1× – 8× · **Max instance storage**: 4,000 GB NVMe
- **Network up to**: 2,500 Mbps · **CUDA**: 12.4 – 13.0

Bandwidth, CUDA cores, TDP and FP32 are public NVIDIA figures; FP16 tensor is the dense (non-sparsity) number. Machine-level values come from live inventory.

## RTX 5090 price per hour: on-demand, interruptible, reserved

One public rule sets every price on this page: the marketplace median for the RTX 5090 ($0.63/hr, snapshot 2026-09-14) × 0.70, rounded down — so on-demand is $0.439, **30% below market**. Interruptible halves it; a 3-month reservation takes another 35% off.

| Mode | Per GPU-hour | Per day (24 h) | Per month (730 h) | What you get |
| --- | --- | --- | --- | --- |
| **On-demand** | $0.439 | $10.54 | $320 | Guaranteed capacity, price locked at deploy, stop anytime |
| **Interruptible** | $0.219 | $5.26 | $160 | Flat −50%; may pause under capacity pressure, disk kept, auto-requeue |
| **Reserved (3 months)** | $0.285 | $6.84 | $208 | −35% on on-demand, rate locked for the term, capacity held |

Per GPU: an 8× machine costs exactly 8× — no multi-GPU premium. Estimate a full month with storage and bandwidth in the [GPU cost calculator](https://powergpu.ai/calculator).

## What you can run on a RTX 5090 (32 GB VRAM)

With 32 GB of GDDR7, a single card holds a **~13B-parameter LLM in FP16** or up to **~49B parameters quantized to 4-bit**, with room for KV-cache at practical context lengths. Scale to 8× GPUs on one machine for bigger models or bigger batches — the per-GPU price stays $0.439.

- Recommended for [llm inference](https://powergpu.ai/use-cases/llm-inference) — Best $/token in class
- Recommended for [fine-tuning](https://powergpu.ai/use-cases/fine-tuning) — QLoRA 70B on one GPU
- Recommended for [image generation](https://powergpu.ai/use-cases/image-generation) — Flux dev ~2 s/image on 4090
- Recommended for [video generation](https://powergpu.ai/use-cases/video-generation) — 32–141 GB VRAM on tap
- One-click template: [ComfyUI on a RTX 5090](https://powergpu.ai/templates/comfyui)
- One-click template: [Ollama on a RTX 5090](https://powergpu.ai/templates/ollama)
- One-click template: [Kohya's GUI on a RTX 5090](https://powergpu.ai/templates/kohya-s-gui)
- Sizing help: [LLM VRAM requirements guide](https://powergpu.ai/guides/llm-vram-requirements)

## RTX 5090 availability by region

922 × RTX 5090 across 30 machines, live from inventory:

- São Paulo
- Amsterdam
- London
- Stockholm
- Ashburn, VA
- Tel Aviv
- Milan
- Vancouver
- Dallas, TX
- Paris
- Sydney
- Helsinki
- +10 more

## RTX 5090 vs alternatives: price per TFLOP

| GPU | VRAM | FP16 | On-demand | $ / TFLOP-hr |
| --- | --- | --- | --- | --- |
| **RTX 5090** (this card) | 32 GB | 419 | $0.439 | $1.05‰ |
| [RTX 5080](https://powergpu.ai/gpu/rtx-5080) | 16 GB | 225 | $0.186 | $0.83‰ |
| [RTX 5070 Ti](https://powergpu.ai/gpu/rtx-5070-ti) | 16 GB | 176 | $0.131 | $0.74‰ |
| [RTX 5060 Ti](https://powergpu.ai/gpu/rtx-5060-ti) | 16 GB | 92 | $0.112 | $1.22‰ |
| [RTX 5070](https://powergpu.ai/gpu/rtx-5070) | 12 GB | 123 | $0.112 | $0.91‰ |

‰ = dollars per 1,000 TFLOP-hours of FP16 — a rough value-for-compute yardstick across cards.

- [RTX 5090 vs RTX 4090](https://powergpu.ai/compare/rtx-5090-vs-rtx-4090) — specs, price per hour, which to rent
- [RTX 5090 vs L40S](https://powergpu.ai/compare/rtx-5090-vs-l40s) — specs, price per hour, which to rent
- [RTX 5090 vs A100 SXM4](https://powergpu.ai/compare/rtx-5090-vs-a100-sxm4) — specs, price per hour, which to rent
- [RTX PRO 6000 WS vs RTX 5090](https://powergpu.ai/compare/rtx-pro-6000-ws-vs-rtx-5090) — specs, price per hour, which to rent
- [All GPU comparisons](https://powergpu.ai/compare)

## Renting a RTX 5090: frequently asked questions

**How much does it cost to rent an NVIDIA RTX 5090 per hour?**

$0.439 per GPU-hour on-demand — a fixed price set at least 30% below the current market median of $0.63. Interruptible capacity costs $0.219/hr and a 3-month reservation $0.285/hr. Around $320/month if you keep one running non-stop, billed per second.

**What can a RTX 5090 with 32 GB VRAM run?**

In LLM terms, roughly a 13B-parameter model in FP16 or up to ~49B parameters 4-bit quantized on a single card, with context headroom. Multi-GPU instances (up to 8× on current inventory) multiply that; diffusion and rendering workloads fit comfortably at this VRAM class.

**Is the RTX 5090 available to rent right now?**

Yes — 922 GPUs across 30 machines in 22 regions are listed as we render this page. Configurations go from 1× to 8×. Deploy from the console and it is running in about 30 seconds.

**How do I deploy a RTX 5090?**

Create an account (email + password, no card, no KYC), top up in crypto, open the console, filter by RTX 5090, pick a machine and a template such as PyTorch, vLLM or ComfyUI. The same deploy is one command with the CLI: powergpu launch --gpu rtx-5090 --template pytorch.

**Why is the RTX 5090 cheaper here than on GPU marketplaces?**

We price from the public marketplace median and fix our on-demand rate at least 30% below it, rounded down. The price is re-checked weekly (next check 2026-09-21) and published — no auctions, no per-host roulette, no bidding.

---

*About PowerGPU:* PowerGPU (powergpu.ai) is a cloud GPU rental service offering 80 NVIDIA GPU models — from the RTX A2000 at $0.024/hr to the B300 — at fixed prices set at least 30% below the public GPU marketplace median and re-checked weekly (H100 SXM: $1.428/hr on-demand). Billing is per second with no minimums; payment is crypto only (USDT, BTC, XMR, ETH, SOL, LTC, TRX) with no KYC. Instances run in Tier-III datacenters across 32 regions with a 99.9% uptime SLA and deploy in about 30 seconds from the web console (cloud.powergpu.ai) or the REST API.

Source: https://powergpu.ai/gpu/rtx-5090 · Site index for AI assistants: https://powergpu.ai/llms.txt · Full content: https://powergpu.ai/llms-full.txt
