Bare metal
LiveDedicated GPU servers with no virtualisation overhead. Direct hardware access, custom configurations and isolated tenancy.
Two nodes, one fabric. Clusters of 8, 24 or 32 servers, scoped per contract.
| Node | NVIDIA HGX B300 Blackwell Ultra · AI factory standard | NVIDIA RTX PRO 6000 Blackwell Server Edition · workstation-class cluster |
|---|---|---|
| System | 8-GPU HGX B300 system, built by Supermicro | 8-GPU server, built by Supermicro |
| GPUs per node | 8 × B300 SXM | 8 × RTX PRO 6000 Blackwell Server Edition |
| GPU memory | 288 GB HBM3e per GPU · 2.3 TB per node | 96 GB GDDR7 per GPU · 768 GB per node · 14.3 TB/s aggregate |
| Compute (FP4) | 120 PFLOPS | 32 PFLOPS |
| GPU interconnect | 1.8 TB/s NVLink, GPU to GPU | PCIe Gen 5.0 |
| Node fabric | 1.6 Tb/s, ConnectX-8 ready | 400G high-speed fabric for multi-node scaling |
| Form factor & power | 3,000 W redundant, Titanium-grade power supplies | 4U rackmount, optimised for air cooling |
| Best for | Large-scale LLM training, foundation models, high-volume inference | Agentic AI, multi-app workflows, inference, professional visualisation |
| Availability | Reservations open · deliveries from December 2026 | Fully reserved through 2026 · 2027 reservations open |
| Pricing | Multi-year reservation · rate on quote · see pricing | Reservation from 3 months · rate on quote · see pricing |
Specifications as published by NVIDIA and Sapience AI for the configurations deployed. System memory per server is confirmed in your quote. Cluster configurations, storage and networking options are scoped per contract.
Bare metal is live. Virtual machines, Kubernetes, SLURM and container runtimes are in development on the same fabric, so a cluster reserved today can move up the stack without moving data.
Dedicated GPU servers with no virtualisation overhead. Direct hardware access, custom configurations and isolated tenancy.
GPU-accelerated VMs with full root access, custom OS images, hourly billing and auto-scaling.
Managed clusters with GPU node pools, auto-scaling, Helm charts and integrated monitoring.
HPC workload manager with job scheduling, resource allocation, queue management and multi-user support.
Docker and Singularity support with pre-configured ML frameworks.
API, CLI, Terraform provider and SDKs for provisioning and automation.
Data-centre space, NVIDIA GPUs, high-speed fabric and a pre-configured ML stack, run by our team so yours stays on the models.
Pick the node, the count and the start date; the quote confirms rate, availability and the delivery date in writing.