NVIDIA GB200 NVL72

Driving the future of accelerated computing

Beyond.pl AI Panel Send an inquiry

One rack.
One massive GPU domain.

The NVIDIA GB200 NVL72 connects 72 Blackwell GPUs and 36 Grace CPUs in a rack-scale, liquid-cooled design.

It boasts a 72-GPU NVIDIA NVLink™ domain that acts as a single, massive GPU and delivers 30x faster real-time trillion-parameter large language model (LLM) inference vs. NVIDIA H100 GPU.

Beyond.pl makes this level of GPU compute available through its enterprise-ready AI Factory, designed for demanding AI workloads. Whether supporting a single large-scale deployment or a long-term AI transformation program, our team brings the expertise needed to develop, deploy and scale AI projects.

Contact us to explore the capabilities of our AI Factory and discuss your requirements with our team.

 

 

Graphics: GB200 motherboard
Source: NVIDIA

The benefits of scaling AI infrastructure with Beyond.pl

Faser access to AI capacity

Launch AI projects without long infrastructure lead times. With 2 MW of GPU capacity available in Q1 2027, an additional 12 MW of high-density, liquid-cooled GPU capacity scheduled for H2 2027, and another 24 MW available in 2028, Beyond.pl provides a clear capacity roadmap for organizations planning large-scale AI deployments and long-term growth.

Designed for large-scale

Built for dedicated AI environments - not one-size-fits-all deployments. Beyond.pl supports large-scale AI projects with access to NVIDIA B300, GB300, GB200 or Vera Rubin platforms, enabling organizations to deploy the architecture that best matches their performance, operational, and growth requirements.

AI infrastructure on your terms

Your AI infrastructure should adapt to your business - not the other way around. We work closely with your team to deliver the right platform, capacity, and deployment model for your workloads, ensuring your AI environment evolves alongside your long-term strategy.

Use cases for NVIDIA GB200 NVL72

Train larger AI models

Accelerate the development of foundation models and large-scale AI architectures with infrastructure designed to handle the highest memory, compute, and networking demands throughout the training lifecycle.

Scale reasoning and agentic AI

Support AI systems that analyse, plan, and evaluate multiple outcomes before responding. GB200 provides the performance required for advanced reasoning, agentic workflows, and compute-intensive inference.

Deliver enterprise AI at scale

Deploy AI applications that serve thousands of concurrent requests while maintaining high throughput and predictable latency. With large GPU memory capacity and rack-scale architecture, GB200 enables efficient production inference for modern AI services.

NVIDIA GB200 NVL72 Specification

Form factor liquid-cooled, rack-scale system
GPU 72 Blackwell Ultra GPUs
36 NVIDIA Grace CPUs
GPU Memory | Bandwidth 13.4 TB HBM3e | 576 TB/s
NVLink Bandwidth 130 TB/s
Fast Memory 30.2 TB
CPU Cores 2,592 Arm Neoverse V2 cores
FP4 Tensor Core Performance 1,440 PFLOPS | 720 PFLOPS
FP8/FP6 Tensor Core Performance 720 PFLOPS

Individual Blackwell Ultra GPU Specification

FP4 Tensor Core per GPU 20 PFLOPS
FP8/FP6 Tensor Core per GPU 10 PFLOPS
INT8 Tensor Core per GPU 10 POPS
FP16/BF16 Tensor Core per GPU 5 PFLOPS
TF32 Tensor Core per GPU 2.5 PFLOPS
FP32 per GPU 80 TFLOPS
FP64/FP64 Tensor Core per GPU 40 TFLOPS
GPU Memory Bandwidth per GPU 186 GB HBM3E | 8 TB/s
Interconnect Fifth-generation NVIDIA NVLink™: 1.8 TB/s
Source: NVIDIA datasheet

The launch of the first commercial AI Factory in CEE

In the second half of 2025, Beyond.pl launched the first commercial AI Factory in the CEE region in collaboration with NVIDIA. The platform was built on NVIDIA DGX SuperPOD reference architecture powered by NVIDIA B200 GPUs, providing enterprise-grade infrastructure for AI training and inference workloads

Read the press release

Explore more NVIDIA AI infrastructure for your next deployment

NVIDIA GB300

Large-scale compute for the next generation of AI models and agentic systems.

The NVIDIA GB300 NVL72 is a fully liquid-cooled, rack-scale system with 72 NVIDIA Blackwell Ultra GPUs and 36 NVIDIA Grace™ CPUs designed for highly complex AI environments that require powerful compute capacity, fast interconnects and efficient workload scaling.

Explore

NVIDIA B300

Next-generation accelerated compute for advanced AI applications.

NVIDIA B300 supports demanding AI workloads across training, fine-tuning and inference. It is designed for organizations that need scalable performance, enterprise-grade reliability and high-performance NVIDIA accelerated computing for large-scale AI deployments.

Explore

NVIDIA B200

High-performance compute for training, fine-tuning and inference of advanced AI models.

NVIDIA B200 is designed for organizations running demanding generative AI, large language model and high-performance computing workloads. It provides access to NVIDIA Blackwell architecture in an integrated enterprise AI system.

Explore

FAQ

What is NVIDIA GB200 NVL72?

NVIDIA GB200 NVL72 is a rack-scale AI platform built on NVIDIA Grace Blackwell architecture. It integrates 72 Blackwell GPUs and 36 Grace CPUs connected by fifth-generation NVLink, designed for large-scale AI training and high-throughput inference workloads. The system uses a hybrid cooling design, with key components such as GPUs and CPUs liquid-cooled, while other components can be air-cooled.

What is the key difference between GB200 NVL72 and GB300 NVL72?

The key difference is that GB300 NVL72 is built on NVIDIA Blackwell Ultra architecture, offering higher GPU memory capacity and improved AI compute performance compared to GB200 NVL72, particularly for large-scale AI training and reasoning inference workloads. GB300 NVL72 also features a fully liquid-cooled rack-scale architecture, while GB200 NVL72 uses a hybrid cooling design, where key components such as GPUs and CPUs are liquid-cooled and other components can be air-cooled.

What is the key difference between GB200 NVL72 and NVIDIA B200?

GB200 NVL72 is a rack-scale system where 72 NVIDIA Blackwell GPUs and 36 Grace CPUs are unified in a single NVLink domain, delivering 130 TB/s of intra-rack bandwidth and enabling the entire rack to operate as one high-bandwidth compute platform. NVIDIA B200 is an individual Blackwell GPU designed to be deployed in server platforms such as 8-GPU B200 systems, which can scale horizontally across servers using high-speed networking. GB200 NVL72 is the right choice for workloads that benefit from tightly integrated, rack-scale GPU computing and maximum intra-rack bandwidth;  
B200-based systems are the right choice for more flexible, server-based deployments. 

Is GB200 NVL72 available as a self-service offering?

No. NVIDIA GB200 is deployed as a tailored infrastructure solution. Because each rack-scale deployment is designed around specific workload and performance requirements, our team will help you determine the right configuration, capacity, and deployment timeline.

How do I get access to GB200 NVL72?

Contact the Beyond.pl technical team to explore the best deployment strategy for your AI initiatives. We’ll help you assess workload demands and build a solution tailored to your requirements.

Ready to run AI at scale?

Deploy demanding AI workloads on NVIDIA GB200 NVL72 infrastructure at Beyond.pl AI Factory. Check available capacity and discuss dedicated compute designed around your AI workloads and long-term requirements.