Deploy high-performance NVIDIA L40S GPU servers optimized for AI training, LLM inference, 3D rendering, and video production. Ada Lovelace with 48 GB of ECC GDDR6 and three NVENC encoders with AV1, fitted to the machine you configure.
The NVIDIA L40S excels in AI training, graphics rendering, video transcoding, and virtualization with breakthrough Ada Lovelace architecture performance.
The L40S GPU achieves remarkable performance metrics: 1466 TFLOPS in Tensor operations, 212 TFLOPS in RT core capabilities, and 91.6 TFLOPS in Single-precision computing power.
Ada Lovelace
48GB GDDR6 with ECC
18,176 pcs.
864 GB/s
350 W
Fourth-generation Tensor Cores with FP8 support deliver outstanding computational performance for AI training and inference workloads.
91.6 teraFLOPS
733 teraFLOPS
1,466 teraFLOPS
212 teraFLOPS
NVIDIA L40S GPU bare metal servers powered by the Ada Lovelace Architecture, optimized for AI training, scientific computing, and high-performance visualization.
The three data-centre cards we fit, compared as the PCIe add-in cards that go in our servers.
| L40S | A100 80 GB | H100 80 GB | |
|---|---|---|---|
| Architecture | Ada Lovelace | NVIDIA Ampere | Hopper |
| Memory | 48GB GDDR6 with ECC | 80GB HBM2e | 80GB HBM2e |
| Memory Bandwidth | 864 GB/s | 1,935 GB/s | 2,039 GB/s |
| Power | Up to 350W | Up to 300W | Up to 350W |
| Bus interface | PCIe Gen 4 x16 | PCIe Gen 4 x16 | PCIe Gen 5 x16 |
| Video encode | 3 NVENC with AV1 | None | None |
| NVLink | No | Connector, bridge by quote | Connector, bridge by quote |
| The card alone, per month | Loading... | Loading... | Loading... |
Also rentable by the month: L40S machines their owners have listed on our marketplace, each at a price its owner set. Some are listed through our sister company, Primcast, and those are rented on primcast.com.
Common questions about deploying and managing NVIDIA L40S GPU-accelerated dedicated servers for AI, rendering, and professional visualization workloads.