Compare / GPU
COMPARE COMPUTE.
Choose up to three GPUs. Suitability depends on the workload — use the reference model to see how each one fits.
Compare / Up to 3 GPUs
3 selected
Precision
Workload
Model fit uses an estimated 158.0 GB at 8K context, batch 1. Blue marks the strongest value in a row — suitability depends on your workload, so there is no universal “best GPU”.
| Spec | NVIDIAL4 | NVIDIARTX 4090 | NVIDIAH100 SXM |
|---|---|---|---|
| VRAM | 24 GB | 24 GB | 80 GB |
| Memory type | GDDR6 | GDDR6X | HBM3 |
| Memory bandwidth | 300 GB/s | 1.01 TB/s | 3.35 TB/s |
| GPU class | Datacenter | Consumer | Datacenter |
| Architecture | Ada Lovelace | Ada Lovelace | Hopper |
| Workload type | InferenceFine-tuningTraining | InferenceFine-tuningTraining | InferenceFine-tuningTraining |
| Model fit · Llama 70B FP16 | 8 × L434.0 GB headroom | 8 × RTX 409034.0 GB headroom | 4 × H100 SXM162.0 GB headroom |
| Power | 72 W | 450 W | 700 W |
| Interconnect | PCIe 4.0 | PCIe 4.0 | NVLink 4 · 900 GB/s |
| Typical use | Power-efficient cloud inference for small models | Developer workstation inference and small-model fine-tuning | Large-scale training and high-throughput serving |
NVIDIA
L4- VRAM
- 24 GB
- Memory type
- GDDR6
- Memory bandwidth
- 300 GB/s
- GPU class
- Datacenter
- Architecture
- Ada Lovelace
- Workload type
- InferenceFine-tuningTraining
- Model fit · Llama 70B FP16
- 8 × L434.0 GB headroom
- Power
- 72 W
- Interconnect
- PCIe 4.0
- Typical use
- Power-efficient cloud inference for small models
NVIDIA
RTX 4090- VRAM
- 24 GB
- Memory type
- GDDR6X
- Memory bandwidth
- 1.01 TB/s
- GPU class
- Consumer
- Architecture
- Ada Lovelace
- Workload type
- InferenceFine-tuningTraining
- Model fit · Llama 70B FP16
- 8 × RTX 409034.0 GB headroom
- Power
- 450 W
- Interconnect
- PCIe 4.0
- Typical use
- Developer workstation inference and small-model fine-tuning
NVIDIA
H100 SXM- VRAM
- 80 GB
- Memory type
- HBM3
- Memory bandwidth
- 3.35 TB/s
- GPU class
- Datacenter
- Architecture
- Hopper
- Workload type
- InferenceFine-tuningTraining
- Model fit · Llama 70B FP16
- 4 × H100 SXM162.0 GB headroom
- Power
- 700 W
- Interconnect
- NVLink 4 · 900 GB/s
- Typical use
- Large-scale training and high-throughput serving