NVIDIA RTX PRO 6000 Blackwell Server Edition in the EUBlackwell for inference, agents and visualisation.

Dedicated clusters of 8, 24 or 32 RTX PRO 6000 servers on a 400G fabric, operated by Sapience in Bratislava. Reserved from three months, single-tenant, under EU law.

Availability
Fully reserved
Through 2026 · reservations for 2027 open
Per server
8 × RTX PRO 6000
96 GB GDDR7 per GPU · 768 GB per server
Specifications

RTX PRO 6000 server specifications

Each server carries eight NVIDIA RTX PRO 6000 Blackwell Server Edition GPUs in a 4U Supermicro chassis optimised for air cooling. Servers are linked on a 400G fabric for multi-node work and delivered as bare metal.

SpecificationNVIDIA RTX PRO 6000 server
System8-GPU server, built by Supermicro
GPUs per server8 × RTX PRO 6000 Blackwell Server Edition
GPU memory96 GB GDDR7 per GPU · 768 GB per server · 14.3 TB/s aggregate
Compute (FP4)32 PFLOPS per server
GPU interconnectPCIe Gen 5.0
Node fabric400G high-speed fabric for multi-node scaling
Form factor4U rackmount, optimised for air cooling
StorageEncrypted NVMe on every node (AES-256 at rest, TLS 1.3 in transit)
TenancySingle-tenant: dedicated compute, storage and network

Specifications as published by NVIDIA and Sapience AI for the configuration deployed. System memory per server is confirmed in your quote.

Cluster sizes

RTX PRO 6000 capacity is reserved by the cluster. Aggregate figures are the per-server specification multiplied by the server count.

ClusterGPUsGPU memoryFP4 compute
8 servers64 × RTX PRO 60006.1 TB256 PFLOPS
24 servers192 × RTX PRO 600018.4 TB768 PFLOPS
32 servers256 × RTX PRO 600024.6 TB1,024 PFLOPS

What fits on 96 GB

As a rule of thumb, model weights take about one byte per parameter at FP8 and two at FP16/BF16. A 70-billion-parameter model at FP8 (about 70 GB of weights) fits on a single RTX PRO 6000 with room for the KV cache; larger models are sharded across the 768 GB of a server. The inference sizing guide works through the numbers.

What it is for

The flexible Blackwell node.

RTX PRO 6000 Blackwell Server Edition is NVIDIA's platform for agentic AI, multi-application workflows and professional visualisation, and a cost-effective node for serving models.

Inference at scale

Serving open-source or proprietary models from dedicated European servers, with a predictable cost per GPU-hour and no shared tenancy.

Inference →

Fine-tuning

Adapter and parameter-efficient fine-tunes of models that fit within 96 GB per GPU or 768 GB per server, on terms from three months.

Fine-tuning →

Agents, pipelines and visualisation

Agent workflows, multi-application pipelines, simulation and professional graphics that mix AI with rendering.

All solutions →

RTX PRO 6000 or HGX B300?

How the two Blackwell nodes compare on memory, interconnect, terms and availability.

Read the comparison →
Questions

RTX PRO 6000 questions.

Anything else: sales@sapienceai.eu or +421 233 329 562.

What is the minimum RTX PRO 6000 reservation?

One cluster of 8 servers (64 GPUs). Terms start at three months, with longer terms on request.

When is RTX PRO 6000 capacity available?

The current capacity is fully reserved through 2026. Reservations for 2027 are open; tell us your start month and we confirm availability in the quote.

How is RTX PRO 6000 priced?

Per GPU-hour on a reserved term, quoted in writing. Rates are in US dollars and exclude VAT.

Is it suitable for multi-node training?

For moderate multi-node jobs, yes, on the 400G fabric. GPUs within a server communicate over PCIe Gen 5.0 rather than NVLink, so large-scale training is better served by HGX B300.

Where does it run?

In a colocation data centre in Bratislava, Slovakia, on systems operated by Sapience. Customer data is processed and stored in the EU.

Reserve RTX PRO 6000 for 2027.

Pick the cluster size, the term and the start month; the quote confirms rate and availability in writing.

Get a quote Talk to our team