Bare metal servers with an NVIDIA A100 (40 or 80 GB) or an H100 80 GB, both as PCIe add-in cards, for AI training, inference and high-performance computing. A machine already racked with its card is handed over in minutes; any other build is fitted or bought in, and dated before you pay.
Enterprise-grade GPU accelerators engineered for AI training, inference, and scientific computing.
Compare technical specifications to select the optimal configuration for your workload requirements.
The A100 GPU delivers exceptional performance, scalability, and reliability for AI training and inference workloads. Built on Ampere architecture with advanced Tensor Cores for accelerated computing at enterprise scale.
Ampere
40 GB HBM2 or 80 GB HBM2e
6,912
1,555 GB/s (40 GB), 1,935 GB/s (80 GB)
250 W (40 GB), 300 W (80 GB)
PCIe Gen 4 x16
None (5 NVDEC decoders)
Hopper, the generation after Ampere, in its PCIe add-in form: the H100 that fits the servers we build. The SXM5 module with HBM3 mounts on an HGX baseboard, and none of our chassis take one, so its figures do not apply here.
Hopper
80 GB HBM2e
2,039 GB/s
350 W
PCIe Gen 5 x16
None (7 NVDEC decoders)
NVIDIA A100 and H100 dedicated servers powered by Ampere and Hopper architectures, optimized for large-scale AI training, LLM inference, and scientific computing applications.
Also rentable by the month: A100 and H100 machines their owners have listed on our marketplace, each at a price its owner set. Some are listed through our sister company, Primcast, and those are rented on primcast.com.
Common questions about deploying and managing enterprise NVIDIA A100 H100 GPU-accelerated dedicated servers for AI training, inference, and high-performance computing.