GPU · NVIDIA · prices checked 2026-09-14
Rent NVIDIA A40 — 48 GB, $0.765/hr on-demand
- VRAM 48 GBGDDR6
- FP16 tensor 144TFLOPS
- PowerScore 101RTX 3090 = 100
- Configs 1–1×PCIe 3.0
- Online now 12 0 regions
Billed per second, price locked at deploy. Storage $0.08/GB/mo · bandwidth $0.01/GB — the whole fee schedule. Next weekly market re-check: 2026-09-21.
GPU · NVIDIA architecture
A cost-efficient card for right-sized jobs: batch inference, smaller models, CI pipelines and experiments where a flagship would idle. Per-second billing makes it perfect for short bursts.
NVIDIA A40 specs: VRAM, TFLOPS, bandwidth
| GPU model | NVIDIA A40 | Architecture | NVIDIA |
|---|---|---|---|
| VRAM | 48 GB GDDR6 | Memory bandwidth | — |
| FP16 tensor perf. | 144 TFLOPS | FP32 perf. | — |
| CUDA cores | — | TDP | — |
| PowerScore (RTX 3090 = 100) | 101 | PCIe generation | Gen 3.0 |
| Multi-GPU | 1× – 1× | Max instance storage | 0 GB NVMe |
| Network up to | 0 Mbps | CUDA | 12.4 – 13.0 |
Bandwidth, CUDA cores, TDP and FP32 are public NVIDIA figures; FP16 tensor is the dense (non-sparsity) number. Machine-level values come from live inventory.
A40 price per hour: on-demand, interruptible, reserved
One public rule sets every price on this page: the marketplace median for the A40 ($1.09/hr, snapshot 2026-09-14) × 0.70, rounded down — so on-demand is $0.765, 30% below market. Interruptible halves it; a 3-month reservation takes another 35% off.
| Mode | Per GPU-hour | Per day (24 h) | Per month (730 h) | What you get |
|---|---|---|---|---|
| On-demand | $0.765 | $18.36 | $558 | Guaranteed capacity, price locked at deploy, stop anytime |
| Interruptible | $0.382 | $9.17 | $279 | Flat −50%; may pause under capacity pressure, disk kept, auto-requeue |
| Reserved (3 months) | $0.497 | $11.93 | $363 | −35% on on-demand, rate locked for the term, capacity held |
Per GPU: an 1× machine costs exactly 1× — no multi-GPU premium. Estimate a full month with storage and bandwidth in the GPU cost calculator.
What you can run on a A40 (48 GB VRAM)
With 48 GB of GDDR6, a single card holds a ~14B-parameter LLM in FP16 or up to ~72B parameters quantized to 4-bit, with room for KV-cache at practical context lengths. Scale to 1× GPUs on one machine for bigger models or bigger batches — the per-GPU price stays $0.765.
- One-click template: Ollama on a A40
- One-click template: SD WebUI Forge on a A40
- One-click template: Whisper WebUI & API on a A40
- Sizing help: LLM VRAM requirements guide
A40 availability by region
12 × A40 across 0 machines, live from inventory:
A40 vs alternatives: price per TFLOP
| GPU | VRAM | FP16 | On-demand | $ / TFLOP-hr |
|---|---|---|---|---|
| A40 this card | 48 GB | 144 | $0.765 | $5.31‰ |
| L40S | 48 GB | 362 | $0.514 | $1.42‰ |
| A800 PCIE | 80 GB | 240 | $0.466 | $1.94‰ |
| A100 PCIE | 80 GB | 312 | $0.374 | $1.20‰ |
| L40 | 48 GB | 181 | $0.235 | $1.30‰ |
‰ = dollars per 1,000 TFLOP-hours of FP16 — a rough value-for-compute yardstick across cards.
Renting a A40: frequently asked questions
How much does it cost to rent an NVIDIA A40 per hour?
$0.765 per GPU-hour on-demand — a fixed price set at least 30% below the current market median of $1.09. Interruptible capacity costs $0.382/hr and a 3-month reservation $0.497/hr. Around $558/month if you keep one running non-stop, billed per second.
What can a A40 with 48 GB VRAM run?
In LLM terms, roughly a 14B-parameter model in FP16 or up to ~72B parameters 4-bit quantized on a single card, with context headroom. Multi-GPU instances (up to 1× on current inventory) multiply that; diffusion and rendering workloads fit comfortably at this VRAM class.
Is the A40 available to rent right now?
Yes — 12 GPUs across 0 machines in 0 regions are listed as we render this page. Configurations go from 1× to 1×. Deploy from the console and it is running in about 30 seconds.
How do I deploy a A40?
Create an account (email + password, no card, no KYC), top up in crypto, open the console, filter by A40, pick a machine and a template such as PyTorch, vLLM or ComfyUI. The same deploy is one command with the CLI: powergpu launch --gpu a40 --template pytorch.
Why is the A40 cheaper here than on GPU marketplaces?
We price from the public marketplace median and fix our on-demand rate at least 30% below it, rounded down. The price is re-checked weekly (next check 2026-09-21) and published — no auctions, no per-host roulette, no bidding.