Data centre networking infrastructure

Documentation

Technical guides for getting your workloads running on AlphaSolutions IT infrastructure.

Getting Started: Requesting your first cluster

Capacity is currently open for reservation. Provisioning timelines are confirmed at contract.

AlphaSolutions IT will GPU clusters on request. There is no self-service console. All cluster requests are handled by our engineering team, who will confirm hardware availability, configure your network environment, and hand over credentials once your cluster is commissioned, to the timeline confirmed in your order form.

To request your first cluster, contact us at [email protected] with the following information: the GPU model you require, the number of nodes, your preferred billing model (on-demand hourly or reserved), and your target start date. If you have specific networking requirements, such as a dedicated VLAN, a private interconnect to an existing environment, or a particular IP range, include those details in your initial message.

Once your request is received, our team will send a formal quote. The quote will specify the hardware configuration, per-hour or monthly rate, and the estimated provisioning timeline. Reserved cluster contracts include a minimum term of six months and are subject to a signed order form before provisioning begins.

After the order is confirmed, we will provision your cluster and send you SSH access credentials, the management IP range, and any relevant network configuration details. For multi-node clusters, we will also provide the InfiniBand or RoCEv2 fabric topology so you can configure your distributed training framework accordingly.

All clusters run on bare-metal hardware. There is no hypervisor layer between your workload and the GPU. You receive root access to each node and full control over the software stack, including the choice of CUDA version, container runtime, and job scheduler. Our engineering team supports customers through provisioning and initial configuration.

Server rack in a data centre aisle
Rack-mounted server chassis hardware

Cluster Specifications: HGX B300 NVL8 reference architecture

The HGX B300 is a multi-GPU compute platform designed for large-scale AI training and inference. Each node contains eight Blackwell Ultra B300 SXM GPUs with 288 GB of HBM3e per GPU, for approximately 2,304 GB (2.3 TB) of pooled GPU memory per node, connected over a fifth-generation NVLink fabric.

Fabric sizing is confirmed per deployment based on cluster size and workload profile.

HGX B300 NVL8 clusters at AlphaSolutions IT are now accepting reservations for Q4 2026 availability in Dubai. Reserved contracts are available from six months. Contact our sales team for configuration options, pricing, and lead times.

Networking: RoCEv2 and InfiniBand topology overview

AlphaSolutions IT clusters support two high-performance interconnect fabrics: RDMA over Converged Ethernet version 2 (RoCEv2) and InfiniBand. The choice of fabric depends on the cluster generation and configuration. H100 SXM clusters are configured with InfiniBand or RoCEv2 depending on cluster generation and customer requirement. HGX B300 clusters use RoCEv2 over 800G Ethernet, which provides high bandwidth at lower per-port cost and greater flexibility for multi-tenant isolation. Final fabric configuration is confirmed at quote.

RoCEv2 transports RDMA operations over standard Ethernet infrastructure using UDP/IP encapsulation. Unlike InfiniBand, which requires a dedicated fabric and subnet manager, RoCEv2 runs on the same Ethernet switches used for general cluster traffic, with Priority Flow Control (PFC) and Explicit Congestion Notification (ECN) enabled on the relevant traffic classes to prevent packet loss and manage congestion.

Each tenant's cluster is isolated at the network layer using dedicated VLANs and separate routing tables.

Structured network cabling and patch panels

Need technical help?

Our engineering team is available to assist with configuration and workload setup.

Contact Sales