Will my model fit on this GPU?
Compute the VRAM you need for a given AI model at a chosen precision and context window, and see exactly which rentable GPUs have the headroom — plus the cheapest provider for each.
Params in memory
30.0B
Dense
Bytes / param
2.0
FP16
Context
131,072
tokens
VRAM required
77 GB
weights + KV + 20% headroom
GPUs that fit
Sorted by VRAM ascending — smallest fitting card first (usually the cheapest).
| GPU | VRAM | FP16 TFLOPS | TDP | Cheapest provider | $/hr | |
|---|---|---|---|---|---|---|
| 80GB | — | — | Lium | $0.65/hr | Open → | |
| 80GB | — | — | io.net | $1.00/hr | Open → | |
| 80GB | — | 300W | no live offers | — | Open → | |
| 80GB | — | — | Clore.ai | $0.99/hr | Open → | |
| 80GB | — | — | Lium | $1.30/hr | Open → | |
| 80GB | — | — | Theta EdgeCloud | $2.29/hr | Open → | |
| 96GB | — | — | Clore.ai | $0.86/hr | Open → | |
| 96GB | — | 1000W | Lambda Labs | $2.29/hr | Open → | |
| 96GB | — | — | no live offers | — | Open → | |
| 96GB | — | — | Vast.ai | $1.20/hr | Open → | |
| 96GB | — | — | Vast.ai | $1.20/hr | Open → | |
| 96GB | — | 600W | Cyfuture AI | $1.04/hr | Open → | |
| 96GB | — | — | Nosana | $0.91/hr | Open → | |
| 96GB | — | 600W | RunPod | $1.69/hr | Open → | |
| 96GB | — | — | Clore.ai | $1.15/hr | Open → | |
| 96GB | — | — | Clore.ai | $0.91/hr | Open → | |
| 96GB | — | — | RunPod | $1.69/hr | Open → | |
| 141GB | — | — | RunPod | $3.59/hr | Open → | |
| 141GB | — | — | Vast.ai | $2.95/hr | Open → | |
| 141GB | — | — | RunPod | $0.50/hr | Open → | |
| 192GB | — | — | DeepInfra | $3.69/hr | Open → | |
| 192GB | — | — | RunPod | $0.50/hr | Open → | |
| 192GB | — | 750W | no live offers | — | Open → | |
| 256GB | — | 1000W | Cyfuture AI | $3.08/hr | Open → | |
| 288GB | — | 1400W | no live offers | — | Open → | |
| 288GB | — | — | io.net | $6.73/hr | Open → | |
| 288GB | — | — | Lium | $8.45/hr | Open → | |
| 288GB | — | 2700W | no live offers | — | Open → |
Ad