NVIDIA · Legacy data center · Accelerator card
NVIDIA Tesla P40 24GB
Exact P40 reference specifications and AI deployment guidance, separated from FusionPointAI commercial inventory.
Verified technical specifications
| Manufacturer | NVIDIA |
|---|---|
| Exact SKU | P40 |
| Product family | NVIDIA P40 |
| Architecture | Pascal |
| Generation | Tesla Legacy |
| Entity type | Accelerator card |
| Lifecycle | Legacy |
| Launch date | The manufacturer has not published this information. |
| VRAM | 24 GB |
| Memory technology | GDDR5 |
| Memory bandwidth | 346 GB/s |
| Power | 250 W |
| Interface | PCIe Gen3 x16 |
| Form factor | Dual slot · PCIe · Accelerator |
| Interconnect | PCIe |
| MIG / partitioning | Not supported |
| Supported numeric formats | FP32, FP16, INT8, INT8, FP4 |
| Software stack | cuda, vllm, sglang, tensorrt-llm |
AI workload suitability
LLM inference
Best suited to smaller checkpoints, compressed models or distributed configurations sized by the recommendation engine.
Fine-tuning and training
Suitable for limited fine-tuning and development when software support, cooling and memory capacity are verified.
Image, video and embeddings
FP16 acceleration supports common generative and vector workloads within the listed VRAM envelope.
Power and cooling
250 W published power requires chassis, airflow and power delivery for this exact form factor.
Multi-GPU use
PCIe is the recorded interconnect path; topology must still be validated at node level.
Software ecosystem
cuda, vllm, sglang, tensorrt-llm
Product timeline
Official manufacturer source
https://nvidianews.nvidia.com/news/new-nvidia-pascal-gpus-accelerate-deep-learning-inference
Last verified: 2026-08-24