NVIDIA · Hopper
NVIDIA H200 SXM 141GB
Best for high-memory LLM inference, quantized 70B+ workloads and batch serving with generous VRAM headroom.
- VRAM
- 141 GB
- Memory
- HBM3e
- Bandwidth
- 4,800 GB/s
USD 2,345/month
~ USD 3.21/hour