Skip to main content

GPU instance

GPU (Graphics Processing Unit) instances provide a GPU-based environment with access to thousands of computing cores. This enables the implementation of various use cases, such as training models on large datasets at high speeds or running high-performance graphics applications. For example, in fields like image recognition, voice recognition, natural language processing, and game development, GPU-based instances offer high levels of performance and processing speed, saving time and cost while ensuring high levels of accuracy. Furthermore, these GPU instances allow for large-scale distributed processing, enabling fast data processing and storage.

  • Applicable types: p5i, p2a, p2i, gn1i, p1i

p5i

Equipped with eight HGX B300 5XM GPUs based on the NVIDIA Blackwell architecture, p5i bare metal instances provide a computing environment optimized for large language model (LLM) training and high-performance inference workloads. These instances provide direct access to bare metal hardware, allowing applications to fully utilize 256 vCPUs and GPU performance. They also include approximately 100TB of enterprise local NVMe SSD storage to eliminate storage I/O bottlenecks when loading large datasets. This hardware architecture delivers strong performance for the following large-scale workloads:

  • Large language model fine-tuning and high-performance inference: The substantial VRAM capacity and computing power of eight B300 GPUs within a single node make these instances suitable for fine-tuning LLMs and running real-time inference services under heavy traffic.
  • High-density single-node AI training: These instances support high-speed training of vision AI, recommendation systems, and natural language processing (NLP) models with limited dependence on communication between nodes.
  • Large-scale data preprocessing and simulation: Approximately 100TB of high-speed local NVMe storage significantly reduces processing time for large data pipelines before training, as well as high-resolution 3D rendering and simulation workloads.
Hardware specifications
  • 6th generation Intel® Xeon® 6 processor with speeds up to 3.9GHz
  • Up to 50Gbps network bandwidth
  • 256 vCPUs, 4,096GiB memory, and 99.84TB local NVMe SSD storage included
  • Up to 8 NVIDIA HGX B300 GPUs
Detailed information
Instance sizeGPUvCPUMemory (GiB)Network bandwidth (Gbps)
p5i.b300.baremetal8256409650

p2a

p2a instances are powered by 3rd generation AMD EPYC 7003 series processors and equipped with NVIDIA A100 Tensor Core GPUs, making them suitable for high-performance computing (HPC) workloads. In Bare Metal Server instances (e.g. p2a.baremetal), applications can directly access the physical resources of the host server, such as processors, memory, and network. Currently, only Bare Metal Server instances are available.

Hardware specifications
  • Up to 3.65GHz 3rd generation AMD EPYC processor (AMD EPYC 7513)
  • Up to 50Gbps network bandwidth
  • Instance sizes supporting up to 128 vCPUs and 1,536GiB memory
  • Up to 8 NVIDIA A100 Tensor Core GPUs
  • Support for AMD instruction set (AVX, AVX2)
Detailed information
Instance sizeGPUvCPUMemory (GiB)Network bandwidth (Gbps)
p2a.baremetal8   128  1536    50

p2i

p2i instances are powered by 3rd generation Intel Xeon Scalable processors and are suitable for general-purpose GPU computing.

Hardware specifications
  • Up to 3.2GHz 3rd generation Intel Xeon Scalable processor (Ice Lake 6338)
  • Up to 50Gbps network bandwidth
  • Instance sizes supporting up to 96 vCPUs and 768GB memory
  • Up to 4 NVIDIA A100 Tensor Core GPUs
  • Support for Intel instruction set (AVX, AVX2, AVX-512)
  • Support for Intel Turbo Boost Technology 2.0
Detailed information
Instance sizeGPUvCPUMemory (GiB)Network bandwidth (Gbps)
p2i.6xlarge124192Max 12.5
p2i.12xlarge248384Max 25
p2i.24xlarge4  96  768   Max 50
NUMA topology

In Non-Uniform Memory Access (NUMA) architecture, each CPU can access its own allocated memory (local memory). NUMA architecture enables high scalability by allowing multiple processors to share memory.

p2i instances have the following NUMA topology depending on the instance size.

Instance sizeNumber of NUMA domainsCores per NUMA domain
p2i.6xlarge1                 12
p2i.12xlarge124
p2i.24xlarge224
NUMA topology architecture
p2i.6xlarge

Image. p2i.6xlarge NUMA topology

p2i.12xlarge

Image. p2i.12xlarge NUMA topology

p2i.24xlarge

Image. p2i.24xlarge NUMA topology

p1i

p1i instances are powered by Gold 5120 Skylake Intel Xeon Scalable processors and are suitable for advanced computational workloads such as machine learning and HPC applications.

Hardware specifications
  • Up to 3.2GHz Gold 5120 Skylake Intel Xeon Scalable processor
  • Up to 50Gbps network bandwidth
  • Instance sizes supporting up to 56 vCPUs and 512GB memory
  • Up to 4 NVIDIA V100 Tensor Core GPUs
  • Support for Intel instruction set (AVX, AVX2, AVX-512)
  • Support for Intel Turbo Boost Technology 2.0
Detailed information
Instance sizeGPUvCPUMemory (GiB)Network bandwidth (Gbps)
p1i.baremetal4  56  512   Up to 50

gn1i

gn1i instances are powered by 2nd generation Intel Xeon Scalable processors and equipped with NVIDIA T4 Tensor Core GPUs, making them suitable for machine learning and graphically intensive workloads.

Hardware specifications
  • Up to 3.9GHz 2nd generation Intel Xeon Scalable processor (Cascade Lake 5220)
  • Up to 50Gpbs network bandwidth
  • Instance sizes supporting up to 64 vCPUs and 256GiB memory
  • Up to 4 NVIDIA T4 Tensor Core GPUs
  • Support for Intel instruction set (AVX, AVX2, AVX-512)
  • Support for Intel Turbo Boost Technology 2.0
Detailed information
Instance sizeGPUvCPUMemory (GiB)Network bandwidth (Gbps)
gn1i.xlarge1416Max 10
gn1i.2xlarge1832Max 10
gn1i.4xlarge11664Max10
gn1i.8xlarge132128Max 25
gn1i.12xlarge448192Max 25
gn1i.16xlarge164256Max 50
NUMA topology

gn1i instances have the following NUMA topology depending on the instance size.

Instance sizeNumber of NUMA domainsCores per NUMA domain
gn1i.xlarge1                 2
gn1i.2xlarge14
gn1i.4xlarge18
gn1i.8xlarge216
gn1i.12xlarge212
gn1i.16xlarge216
NUMA topology architecture
gn1i.xlarge

Image. gn1i.xlarge NUMA topology

gn1i.2xlarge

Image. gn1i.2xlarge NUMA topology

gn1i.4xlarge

Image. gn1i.4xlarge NUMA topology

gn1i.8xlarge

Image. gn1i.8xlarge NUMA topology

gn1i.12xlarge

Image. gn1i.12xlarge NUMA topology

gn1i.16xlarge

Image. gn1i.16xlarge NUMA topology