NVIDIA A100 SXM 80GB GPU accelerator

NVIDIA · Ampere

NVIDIA A100 SXM 80GB

Best for proven CUDA workloads, LoRA fine-tuning, research and stable production inference.

Monthly rental

USD 915/month

~ USD 1.25/hour based on a 730-hour month

Availability: Available

Compare GPUs

Technical specifications

Identity

Vendor NVIDIA
Model NVIDIA A100 SXM 80GB
SKU A100 SXM 80GB
Architecture Ampere
Hardware class datacenter gpu

Memory

VRAM 80 GB
Memory type HBM2e
Memory bandwidth 2039 GB/s

Compute

Precision Dense With sparsity
FP64 9.7 TFLOPS
FP64 Tensor 19.5 TFLOPS
FP32 19.5 TFLOPS
TF32 Tensor 156 TFLOPS 312 TFLOPS
BF16 Tensor 312 TFLOPS 624 TFLOPS
FP16 Tensor 312 TFLOPS 624 TFLOPS
INT8 Tensor 624 TOPS 1248 TOPS

Platform

Form factor / interface SXM module
Interface PCIe Gen4 x16 host interface
TDP 400 W
NVLink 600 GB/s NVLink
NVSwitch Supported
MIG Up to 7 MIGs @ 10 GB
Max GPUs per node 8

Software

CUDA Supported
Verification OFFICIAL_VENDOR
Source www.nvidia.com

Workload fit

  • LLM inferenceVery Good
  • LLM fine-tuningVery Good
  • Full model trainingVery Good
  • Image / video generationVery Good
  • Embeddings / rerankingExcellent

Frameworks

  • PyTorch
  • TensorFlow
  • JAX
  • Hugging Face
  • vLLM
  • TensorRT
  • ONNX Runtime
  • CUDA

Model compatibility guidance

7B/8B inference

Recommended

13B/14B inference

Recommended with quantisation or FP16/BF16 headroom review

30B/32B inference

Recommended with quantisation for production headroom

65B/70B inference

Recommended for quantised inference; review context and concurrency

Fine-tuning

Suitable for LoRA/QLoRA; full fine-tuning depends on model size and topology

Large training

Use multi-GPU nodes with high-bandwidth topology

Related GPUs

AI GPU hosting FAQ

Is NVIDIA A100 SXM 80GB suitable for LLM inference?

NVIDIA A100 SXM 80GB is rated Very Good for inference in the FusionPointAI catalogue. Final sizing still depends on model size, precision, context length and concurrency.

How is monthly GPU pricing calculated?

The public monthly price is the authoritative price version. The displayed hourly equivalent is derived from that monthly price using a 730-hour month.

Can I rent more than one GPU?

Multi-GPU configurations are supported when inventory and node topology allow it. This SKU currently lists up to 8 GPUs per node.