Servers · GPU AI · Rack

Rack GPU AI servers — 2 to 6 GPUs per node

Thermally validated, power-balanced, and prepared with driver and firmware pinning for stable AI training and inference. PCIe Gen4/Gen5, NVLink on select GPUs, U.2/U.3 NVMe and up to 200G fabric.

About rack GPU AI servers

Density you can add to, a node or a rack at a time

These nodes deliver GPU density with enterprise thermal, power and management behaviour behind it. They are built for AI workloads at scale, and sized so a programme can start small and expand by node or by rack rather than by replacement.

Every server is validated under burn-in with the GPUs at peak draw. A GPU node that passes a short benchmark and then throttles in week three is the common failure, and it is a thermal and power design problem rather than a component one.

Pick the density that matches the work: 2 GPUs for inference and fine-tuning, 4 for mainstream training, 6 where throughput per rack unit is what the budget is buying.

Three densities

Match the node to the workload

Sizing examples

Three profiles, end to end

Starting points that show how the whole node moves when GPU count changes — CPU, memory, storage and fabric all follow.

Illustrative. Final sizing follows the framework, model size and batch size you actually run.

Interconnect

NVLink or PCIe — which do you need?

We map your framework and batch size to the right interconnect. Many deployments run excellently on PCIe alone — NVLink is a requirement, not an upgrade.

Feature highlights

What holds the clocks up in month six

A GPU node is bought on peak numbers and lived with on sustained ones. These are the parts of the design that decide whether the second figure resembles the first.

Thermals, power rails and NVMe separation do not appear on a spec sheet comparison, and they are the three things that decide whether a training run finishes at the throughput it started with.

Deployment notes

What to plan for around the node

As much of a GPU deployment is decided outside the chassis as inside it. These are the parts to settle early.

At a glance

The rack GPU envelope

Size it first

Work the numbers before you specify

The calculators run the same arithmetic our engineers do, so you can arrive with a starting configuration rather than a blank page.

Related range

Related platforms

Quieter, or at the desk.