Powered by eight NVIDIA L40 GPUs, each with 48 GB of GDDR6 memory, these instances are designed for exceptional versatility across AI and visual computing. Connected via PCIe, they offer a balanced profile for workloads that include AI inference, rendering, and video processing. They are an ideal entry point for deploying AI-powered services that also have a significant graphics component.
Specifications
1 Usable storage is less than the raw capacity when configured as RAID. See RAID layout and throughput to learn about RAID configurations and their performance characteristics.
Primary use cases
General-purpose AI inference, small-scale model fine-tuning, high-resolution rendering, and virtual desktop infrastructure (VDI).
Recommended models
Inference for models up to ~40B parameters like Gemma 4 26B or Qwen 3.6 27B; efficient fine-tuning of smaller sub-15B models. Last modified on June 23, 2026