Skip to main content

NVIDIA · Platform · Superchip

NVIDIA GH200 Grace Hopper Superchip 96GB

Exact GH200 96GB reference specifications and AI deployment guidance, separated from FusionPointAI commercial inventory.

Recent generation Reference only
Neutral technical reference illustration for NVIDIA GH200 Grace Hopper Superchip 96GB; not a product photograph
Reference illustration; consult the verified specification table for exact hardware identity. This is not a product photograph.

Verified technical specifications

ManufacturerNVIDIA
Exact SKUGH200 96GB
Product familyNVIDIA GH200
ArchitectureHopper
GenerationGrace Hopper
Entity typeSuperchip
LifecycleRecent generation
Launch dateThe manufacturer has not published this information.
VRAM96 GB
Memory technologyHBM3
Memory bandwidth4,000 GB/s
Power1000 W
InterfaceNVLink-C2C
Form factorCPU-GPU · Superchip
InterconnectNVLink-C2C
MIG / partitioningSupported
AI precision supportFP64, FP32, TF32, BF16, FP16, FP8, INT8
Software stackcuda, vllm, sglang, tensorrt-llm

AI workload suitability

LLM inference

High memory capacity supports large checkpoints and multi-GPU inference when runtime and topology are certified.

Fine-tuning and training

Suitable for limited fine-tuning and development when software support, cooling and memory capacity are verified.

Image, video and embeddings

FP16 acceleration supports common generative and vector workloads within the listed VRAM envelope.

Power and cooling

1000 W published power requires chassis, airflow and power delivery for this exact form factor.

Multi-GPU use

NVLink-C2C is the recorded interconnect path; topology must still be validated at node level.

Software ecosystem

cuda, vllm, sglang, tensorrt-llm

Product timeline

nvidia-gh200-96gb @ 45ceb776e39ae795b323418d5ab43f7423707f3e27bc97dc1cff6c9c5b862b01

Official source

Media gallery

Official manufacturer source

https://www.nvidia.com/en-us/data-center/grace-hopper-superchip/

Last verified: 2026-08-24