Louer Nvidia GeForce RTX 4090.
Daily median across providers.
All providers carrying this GPU.
Should you rent or own?
Suitable workloads.
Run these on this GPU.
- Gemma 3 27B 1× · int4
- Llama 3.2 11B Vision 1× · fp16
- Qwen 2.5 Coder 32B 1× · fp16
Hyperscaler bundles.
Frequently asked.
What's a cloud instance bundle for the Nvidia GeForce RTX 4090?
Why is per-GPU pricing on instances different from raw GPU rental rates?
Which hyperscaler is cheapest for the Nvidia GeForce RTX 4090?
Can I run my own image or container on these instances?
How do I get the cheapest rate on the Nvidia GeForce RTX 4090 overall?
Pre-configured instances on hyperscalers.
Whole-instance bundles (GPU + vCPU + RAM + disk) on the major clouds. Per-GPU rate often drops as the count rises. View = spec page · Launch = sign up (affiliate).
| Provider | Instance | GPUs | vCPU | RAM | Disk | $/hr | $/hr per GPU | |
|---|---|---|---|---|---|---|---|---|
| RTX4090x8_2 | 8× | 88 | 228 GB | 7000 GB | $3.52/hr | $0.44/hr | ||
| RTX4090x8_2 | 8× | 88 | 228 GB | 7000 GB | $3.52/hr | $0.44/hr | ||
| RTX4090x8_2 | 8× | 88 | 228 GB | 7000 GB | $3.52/hr | $0.44/hr | ||
| 8x RTX4090_PCIE | 8× | 88 | 228 GB | 7000 GB | $3.84/hr | $0.48/hr | ||
| 8x RTX4090_PCIE | 8× | 88 | 228 GB | 7000 GB | $3.84/hr | $0.48/hr |
Cloud instances — common questions.
What's a cloud instance bundle for the Nvidia GeForce RTX 4090?
Why is per-GPU pricing on instances different from raw GPU rental rates?
Which hyperscaler is cheapest for the Nvidia GeForce RTX 4090?
Can I run my own image or container on these instances?
How do I get the cheapest rate on the Nvidia GeForce RTX 4090 overall?
Models that run on this GPU.
GPU-count + quantization recommendations covering fine-tuning, inference, and run-it-yourself scenarios on the Nvidia GeForce RTX 4090.
Gemma 3 27B
Snug fit, recommended for hobbyist self-hosting
Llama 3.2 11B Vision
Qwen 2.5 Coder 32B
24GB just enough — try int8 if OOM
DeepSeek R1 Distill Qwen 14B
Fits with room for context.
DeepSeek R1 Distill Qwen 7B
Whisper Large v3
Real-time transcription with headroom.
FLUX.1 Dev
Sweet spot for FLUX.1 Dev.
FLUX.1 Schnell
<2s per image with 4-step distillation.
Stable Diffusion 3.5 Large
Stable Diffusion 3.5 Medium
Llama 3.1 8B
Comfortable single-GPU local target.
Mistral 7B v0.3
TheDrummer: Skyfall 36B V2
Stable Diffusion XL
Mistral 7B v0.2
Qwen3.6 35B A3b Fp8
Mixtral 8x22B
Qwen 2.5 72B
Qwen: Qwen3.5-35B-A3B
Llama 3.3 70B
Quantized only — consumer multi-GPU
DeepSeek R1 Distill Qwen 32B
FLUX.1 Dev
Halves VRAM, ~95% quality.
Stable Diffusion 3.5 Large
FP8 cuts VRAM by half.