Salad
Distributed inference cloud — RTX 3090 $0.09/h, RTX 4090 $0.16/h
- Cheapest consumer GPUs — RTX 3090 from $0.09/h
- Massive horizontal scale (1000+ nodes)
RTX 4090 cloud comparison · April 2026
The consumer-tier sweet spot — 5 clouds offer RTX 4090 (24GB) from $0.13/h. Best value for SDXL/FLUX, Llama 3 8B and hobby ML.
The RTX 4090 is the consumer-tier sweet spot for AI/ML in 2026. With 24 GB of GDDR6X VRAM and Ada Lovelace architecture, it punches well above its consumer-card weight class — running Llama 3 8B, Stable Diffusion 3 / FLUX, Mistral 7B and most ≤13B models comfortably.
5 GPU clouds offer 4090 instances. Pricing spans $0.13/h to $0.74/h for the same card. The marketplace clouds (Vast.ai, RunPod Community) offer 4090s at consumer-rental prices because providers monetize idle gaming hardware.
For hobbyists, indie devs and Stable Diffusion enthusiasts, the RTX 4090 is the highest-value GPU on the market — paying for an A100 for these workloads usually doesn't pay back.
None of them rent a 4090. NVIDIA's licence terms keep GeForce cards out of most datacentres, so the hyperscalers stock datacentre parts instead — AWS, Azure and Google Cloud all list A100 and H100, and Lambda Labs runs A100, H100 and the Quadro RTX 6000. If you specifically need a 4090, a marketplace is the only route. If you just need 24 GB of VRAM, an L4 or an A10 on a big cloud is the boring alternative.
Vast.ai's $0.13/h is the low end of a live marketplace, not a rate you can count on. Vast's own distribution for the 4090 puts the median at $0.36/h and the 90th percentile at $0.53/h, so a machine grabbed at a busy moment costs roughly 4x the headline. RunPod Community, at a flat $0.34/h, is within a couple of cents of the Vast median without the variance.
Per unit of work the 4090 is the cheaper card whenever the job fits in 24 GB: at $0.34/h and 165 TFLOPS FP16 it runs about $0.002 per TFLOP-hour, against roughly $0.0035 for an A100 80GB at $1.19/h and 312 TFLOPS. The A100 wins the moment your model no longer fits and you would otherwise be splitting it across cards.
| Provider | Starting Price | Top GPUs | Highlights | Rating | CTA |
|---|---|---|---|---|---|
| Salad | from $0.02/h | RTX 3090, RTX 4090, RTX 3080 ≤24GB |
| ★★★★☆ | View pricing |
| Vast.ai Editor's Choice | from $0.03/h | RTX 3090, RTX 4090, A100 ≤80GB |
| ★★★★☆ | View pricing |
| TensorDock | from $0.10/h | RTX 4090, RTX 3090, A100 80GB ≤80GB |
| ★★★★☆ | View pricing |
| RunPod Editor's Choice | from $0.16/h | RTX A5000, RTX 3090, RTX 4090 ≤80GB |
| ★★★★★ | View pricing |
| Novita AI | from $0.33/h | RTX 4090, RTX 5090, RTX 6000 Ada ≤48GB |
| ★★★★☆ | View pricing |
Distributed inference cloud — RTX 3090 $0.09/h, RTX 4090 $0.16/h
Cheapest GPU cloud — peer-to-peer marketplace for budget training
Marketplace GPU cloud — RTX A4000 from $0.10/h, RTX 4090 $0.35/h, H100 SXM5 $2.25/h
Best value GPU cloud — huge selection, community + secure cloud
Consumer GPUs by the hour, spot prices from $0.17/h
Vast.ai starts at $0.13/h, but that is the marketplace floor — the median 4090 there is $0.36/h and the 90th percentile $0.53/h, so plan against the median. RunPod Community is a flat $0.34/h with better reliability, TensorDock $0.35/h. For production, RunPod Secure 4090 is $0.74/h. Verified August 2026.
Llama 3 8B fine-tuning fits comfortably on a single 4090 (24 GB) with QLoRA. Llama 3 70B does NOT fit — you need ≥40 GB VRAM (A100 40GB or H100). For full Llama 3 8B fine-tuning: 2× 4090 with FSDP works but is slower than a single A100.
For SD/SDXL/FLUX inference and LoRA training, the 4090 is faster (Ada Lovelace + higher clock) and ~3× cheaper than A100 40GB. A100 only wins for very large batches or training on millions of images. Default to 4090 for image AI.
Break-even point: a $1,500 RTX 4090 pays for itself vs $0.34/h cloud rental at ~4,400 hours of usage (~6 months 24/7). For sporadic use, renting wins; for always-on workloads, buying wins. Add electricity (~$300/year at 24/7) to the buy side.
Vast.ai and RunPod Community are marketplaces where individual hosts set their own rates, so supply and demand move the price. Vast.ai's published range for the 4090 spans $0.13 to $2.72/h against a $0.36/h median. Sort by price at the moment you rent rather than trusting any number you read in an article, including this one.
Rarely, and not reliably. The 4090 has no NVLink, so multi-GPU scaling depends on PCIe bandwidth and falls off faster than it does on SXM A100 or H100 nodes. Multi-GPU 4090 boxes do exist on Vast.ai and TensorDock, but for distributed training the interconnect becomes the bottleneck long before the card does.
Community 4090s on Vast.ai and RunPod Community use consumer hardware with variable uptime — fine for batch jobs and tinkering, not for production APIs. RunPod Secure 4090s run in datacenter-class facilities with uptime SLAs.
Get an email when GPU prices drop or availability changes at your preferred provider.
No spam. Unsubscribe any time.