NVIDIA · Data center · Module
NVIDIA A100 SXM4 40GB
Exact A100 SXM4 40GB reference specifications and AI deployment guidance, separated from FusionPointAI commercial inventory.
Verified technical specifications
| Manufacturer | NVIDIA |
|---|---|
| Exact SKU | A100 SXM4 40GB |
| Product family | NVIDIA A100 |
| Architecture | Ampere |
| Generation | Ampere Data Center |
| Entity type | Module |
| Lifecycle | Previous generation |
| Launch date | The manufacturer has not published this information. |
| VRAM | 40 GB |
| Memory technology | HBM2 |
| Memory bandwidth | 1,555 GB/s |
| Power | 400 W |
| Interface | PCIe Gen4 host interface |
| Form factor | SXM4 · Module |
| Interconnect | NVLink |
| MIG / partitioning | Supported |
| AI precision support | FP64, FP32, TF32, BF16, FP16, INT8, INT4 |
| Software stack | cuda, vllm, sglang, tensorrt-llm |
AI workload suitability
LLM inference
Best suited to smaller checkpoints, compressed models or distributed configurations sized by the recommendation engine.
Fine-tuning and training
Designed for sustained accelerator workloads; training suitability depends on interconnect, precision and cluster topology.
Image, video and embeddings
FP16 acceleration supports common generative and vector workloads within the listed VRAM envelope.
Power and cooling
400 W published power requires chassis, airflow and power delivery for this exact form factor.
Multi-GPU use
NVLink is the recorded interconnect path; topology must still be validated at node level.
Software ecosystem
cuda, vllm, sglang, tensorrt-llm
Product timeline
Official manufacturer source
https://www.nvidia.com/en-us/data-center/a100/
Last verified: 2026-08-24