NVIDIA · Data center · Accelerator card
NVIDIA A30 24GB
Exact A30 reference specifications and AI deployment guidance, separated from FusionPointAI commercial inventory.
Verified technical specifications
| Manufacturer | NVIDIA |
|---|---|
| Exact SKU | A30 |
| Product family | NVIDIA A30 |
| Architecture | Ampere |
| Generation | Ampere Data Center |
| Entity type | Accelerator card |
| Lifecycle | Previous generation |
| Launch date | The manufacturer has not published this information. |
| VRAM | 24 GB |
| Memory technology | HBM2 |
| Memory bandwidth | 933 GB/s |
| Power | 165 W |
| Interface | PCIe Gen4 x16 |
| Form factor | Dual slot · PCIe · Accelerator |
| Interconnect | NVLink |
| MIG / partitioning | Supported |
| Supported numeric formats | FP64, FP32, TF32, BF16, FP16, INT8, INT4 |
| Software stack | cuda, vllm, sglang, tensorrt-llm |
AI workload suitability
LLM inference
Best suited to smaller checkpoints, compressed models or distributed configurations sized by the recommendation engine.
Fine-tuning and training
Designed for sustained accelerator workloads; training suitability depends on interconnect, precision and cluster topology.
Image, video and embeddings
FP16 acceleration supports common generative and vector workloads within the listed VRAM envelope.
Power and cooling
165 W published power requires chassis, airflow and power delivery for this exact form factor.
Multi-GPU use
NVLink is the recorded interconnect path; topology must still be validated at node level.
Software ecosystem
cuda, vllm, sglang, tensorrt-llm
Product timeline
Official manufacturer source
https://www.nvidia.com/en-us/data-center/products/a30-gpu/
Last verified: 2026-08-24