Skip to main content
Powered by NVIDIA GB200 GPUs with InfiniBand connectivity, these instances deliver a monumental leap in performance over the previous generation, setting the new standard for large-scale AI and high-performance computing. The GB200 instances are equipped with 13.4 TB of high-bandwidth GPU memory per rack and fifth-generation NVLink with up to 130 TB/s total bandwidth. Blackwell GPUs feature a second-generation Transformer Engine with new FP4 and FP6 precision, drastically accelerating LLM inference and training. The NVL72 architecture provides unprecedented memory capacity and ultra-fast GPU-to-GPU communication for distributed computing at the largest scale.

Specifications

1 Usable storage is less than the raw capacity when configured as RAID. See RAID layout and throughput to learn about RAID configurations and their performance characteristics.

Primary use cases

Training foundation models from scratch (500B+ parameters), massive-scale inference, and developing frontier AI systems. DeepSeek V4 Pro, Kimi K2.6, GLM 5.1, and other future advanced GPT-5 class models, complex Mixture-of-Experts (MoE) models, and large multimodal systems built for coding and agentic workflows.
Last modified on June 23, 2026