GPU cloud catalogue

Compare public GPU rental options by VRAM, architecture, availability, monthly price and AI workload fit.

NVIDIA H200 SXM 141GB GPU accelerator
NVIDIA · Hopper

NVIDIA H200 SXM 141GB

Best for high-memory LLM inference, quantized 70B+ workloads and batch serving with generous VRAM headroom.

VRAM
141 GB
Memory
HBM3e
Bandwidth
4,800 GB/s
Limited availability Best for LLM inference · LLM fine-tuning
USD 2,345/month ~ USD 3.21/hour
NVIDIA H100 SXM 80GB GPU accelerator
NVIDIA · Hopper

NVIDIA H100 SXM 80GB

Best for production LLM serving, 70B-class quantized inference and high-bandwidth multi-GPU training.

VRAM
80 GB
Memory
HBM3
Bandwidth
3,350 GB/s
Available Best for LLM inference · LLM fine-tuning
USD 1,860/month ~ USD 2.55/hour
NVIDIA H100 PCIe 80GB GPU accelerator
NVIDIA · Hopper

NVIDIA H100 PCIe 80GB

Best for powerful single-node inference, smaller fine-tuning jobs and cost-efficient Hopper access.

VRAM
80 GB
Memory
HBM2e
Bandwidth
2,000 GB/s
Limited availability Best for LLM inference · LLM fine-tuning
USD 1,579/month ~ USD 2.16/hour
NVIDIA A100 SXM 80GB GPU accelerator
NVIDIA · Ampere

NVIDIA A100 SXM 80GB

Best for proven CUDA workloads, LoRA fine-tuning, research and stable production inference.

VRAM
80 GB
Memory
HBM2e
Bandwidth
2,039 GB/s
Available Best for LLM inference · LLM fine-tuning
USD 915/month ~ USD 1.25/hour
NVIDIA L40S 48GB GPU accelerator
NVIDIA · Ada Lovelace

NVIDIA L40S 48GB

Best for image generation, computer vision, embeddings and mid-size LLM inference.

VRAM
48 GB
Memory
GDDR6 ECC
Bandwidth
864 GB/s
Available Best for LLM inference · Image / video generation
USD 506/month ~ USD 0.69/hour
NVIDIA L4 24GB GPU accelerator
NVIDIA · Ada Lovelace

NVIDIA L4 24GB

Best for small-model inference, embeddings, reranking and lightweight production workloads.

VRAM
24 GB
Memory
GDDR6
Bandwidth
300 GB/s
Available Best for Embeddings / reranking
USD 250/month ~ USD 0.34/hour