NVIDIA · Legacy data center · Module
NVIDIA Tesla V100 SXM2 16GB
Exact V100 SXM2 16GB reference specifications and AI deployment guidance, separated from FusionPointAI commercial inventory.
Verified technical specifications
| Manufacturer | NVIDIA |
|---|---|
| Exact SKU | V100 SXM2 16GB |
| Product family | NVIDIA V100 |
| Architecture | Volta |
| Generation | Tesla Legacy |
| Entity type | Module |
| Lifecycle | Legacy |
| Launch date | The manufacturer has not published this information. |
| VRAM | 16 GB |
| Memory technology | HBM2 |
| Memory bandwidth | 900 GB/s |
| Power | 300 W |
| Interface | SXM2 module interface |
| Form factor | SXM2 · Module |
| Interconnect | NVLink |
| MIG / partitioning | Not supported |
| Supported numeric formats | FP64, FP32, FP16, INT8, FP4 |
| Software stack | cuda, vllm, sglang, tensorrt-llm |
AI workload suitability
LLM inference
Best suited to smaller checkpoints, compressed models or distributed configurations sized by the recommendation engine.
Fine-tuning and training
Suitable for limited fine-tuning and development when software support, cooling and memory capacity are verified.
Image, video and embeddings
FP16 acceleration supports common generative and vector workloads within the listed VRAM envelope.
Power and cooling
300 W published power requires chassis, airflow and power delivery for this exact form factor.
Multi-GPU use
NVLink is the recorded interconnect path; topology must still be validated at node level.
Software ecosystem
cuda, vllm, sglang, tensorrt-llm
Product timeline
Official manufacturer source
Last verified: 2026-08-24