按模型 sizing 按库存报价 版本化报价

面向 AI 的高性能 GPU 云

租用数据库定价 GPU,用于 LLM 推理、微调、多模态工作负载、计算机视觉和研究,并查看清晰规格。

6
公开 GPU 选项
6
已定价租赁 GPU
730
小时/月基准

GPU 目录与可用容量

Browse hardware the calculator understands before running a workload-specific recommendation. Catalogue entries are separate from live capacity and approved pricing.

NVIDIA H200 SXM 141GB GPU accelerator
NVIDIA · Hopper

NVIDIA H200 SXM 141GB

适合大模型、长上下文、重型微调和高显存 AI 工作负载。

显存
141 GB
内存
HBM3e
带宽
4,800 GB/s
Limited capacity 适合 工作负载评估
USD 2,345/月 ~ USD 3.21/小时
NVIDIA H100 SXM 80GB GPU accelerator
NVIDIA · Hopper

NVIDIA H100 SXM 80GB

适合量化 70B 推理、LoRA 微调和高吞吐服务。

显存
80 GB
内存
HBM3
带宽
3,350 GB/s
Available 适合 工作负载评估
USD 1,860/月 ~ USD 2.55/小时
NVIDIA H100 PCIe 80GB GPU accelerator
NVIDIA · Hopper

NVIDIA H100 PCIe 80GB

适合量化 70B 推理、LoRA 微调和高吞吐服务。

显存
80 GB
内存
HBM2e
带宽
2,000 GB/s
Limited capacity 适合 工作负载评估
USD 1,579/月 ~ USD 2.16/小时
NVIDIA A100 SXM 80GB GPU accelerator
NVIDIA · Ampere

NVIDIA A100 SXM 80GB

适合量化 70B 推理、LoRA 微调和高吞吐服务。

显存
80 GB
内存
HBM2e
带宽
2,039 GB/s
Available 适合 工作负载评估
USD 915/月 ~ USD 1.25/小时
NVIDIA L40S 48GB GPU accelerator
NVIDIA · Ada Lovelace

NVIDIA L40S 48GB

适合中型 LLM、图像生成、计算机视觉和成本可控的生产推理。

显存
48 GB
内存
GDDR6 ECC
带宽
864 GB/s
Available 适合 工作负载评估
USD 506/月 ~ USD 0.69/小时
NVIDIA L4 24GB GPU accelerator
NVIDIA · Ada Lovelace

NVIDIA L4 24GB

适合小模型推理、embedding、reranking、开发和成本敏感 AI 服务。

显存
24 GB
内存
GDDR6
带宽
300 GB/s
Available 适合 工作负载评估
USD 250/月 ~ USD 0.34/小时

您想运行什么?

简单模式
高级模式

从已验证 GPU 库存中选择

先用目录选择硬件;当模型规模、上下文或预算很关键时,再使用推荐器。

NVIDIA H200 SXM 141GB GPU accelerator

NVIDIA H200 SXM 141GB

适合大模型、长上下文、重型微调和高显存 AI 工作负载。

显存
141 GB
架构
Hopper
价格
USD 2,345
查看 GPU

GPU 规模估算方式

计算器会先估算模型权重、KV 缓存、运行时工作区、工作负载开销和生产余量,再筛选库存。

推理与训练

推理和训练分别估算。训练会包含优化器状态、梯度或适配器、激活、运行空间和拓扑限制。

精度与量化

较低精度可以减少权重内存,但量化元数据、运行空间和技术限制仍然需要考虑。

版本化报价逻辑

报价请求会绑定模型快照、计算器版本、所选配置和价格版本。

AI GPU hosting 常见问题

How is VRAM estimated?

The estimator combines model weights, KV cache, runtime workspace, batching, workload overhead and safety headroom.

Is training sized differently from inference?

Yes. Fine-tuning and training include gradients, optimiser state, activations and topology constraints.

Can I use a custom model?

Yes. You can enter parameter count and architecture details; lower-confidence estimates are labelled.

Are recommendations benchmarks?

No. Estimated results and measured benchmark data are labelled separately.

Can unavailable GPUs be purchased?

No. Unavailable hardware can remain in the catalogue, but it is not offered as immediately purchasable inventory.

How does pricing work?

Approved prices are shown where cost and market inputs have been reviewed. Otherwise the offer remains request-quote.