7B/8B inference
Recommended
NVIDIA · Hopper
Best for powerful single-node inference, smaller fine-tuning jobs and cost-efficient Hopper access.
Monthly rental
~ USD 2.16/hour based on a 730-hour month
Availability: Limited availability
Compare GPUs| Vendor | NVIDIA |
|---|---|
| Model | NVIDIA H100 PCIe 80GB |
| SKU | H100 PCIe |
| Architecture | Hopper |
| Hardware class | datacenter gpu |
| VRAM | 80 GB |
|---|---|
| Memory type | HBM2e |
| Memory bandwidth | 2000 GB/s |
| Precision | Dense | With sparsity |
|---|---|---|
| FP64 | 26 TFLOPS | — |
| FP64 Tensor | 51 TFLOPS | — |
| FP32 | 51 TFLOPS | — |
| TF32 Tensor | 378 TFLOPS | 756 TFLOPS |
| BF16 Tensor | 756.5 TFLOPS | 1513 TFLOPS |
| FP16 Tensor | 756.5 TFLOPS | 1513 TFLOPS |
| FP8 Tensor | 1513 TFLOPS | 3026 TFLOPS |
| INT8 Tensor | 1513 TOPS | 3026 TOPS |
| Form factor / interface | PCIe dual-slot accelerator card |
|---|---|
| Interface | PCIe Gen5 x16 |
| TDP | 350 W |
| NVLink | 600 GB/s NVLink Bridge |
| NVSwitch | Not supported |
| MIG | Up to 7 MIGs @ 10 GB |
| Max GPUs per node | 8 |
| CUDA | Supported |
|---|---|
| Verification | OFFICIAL_VENDOR |
| Source | www.nvidia.com |
Recommended
Recommended with quantisation or FP16/BF16 headroom review
Recommended with quantisation for production headroom
Recommended for quantised inference; review context and concurrency
Suitable for LoRA/QLoRA; full fine-tuning depends on model size and topology
Use multi-GPU nodes with high-bandwidth topology
80 GB HBM3 · Available
USD 1,860/month
80 GB HBM2e · Available
USD 915/month
48 GB GDDR6 ECC · Available
USD 506/month
NVIDIA H100 PCIe 80GB is rated Very Good for inference in the FusionPointAI catalogue. Final sizing still depends on model size, precision, context length and concurrency.
The public monthly price is the authoritative price version. The displayed hourly equivalent is derived from that monthly price using a 730-hour month.
Multi-GPU configurations are supported when inventory and node topology allow it. This SKU currently lists up to 8 GPUs per node.