Dynamic GPU slicing with Red Hat OpenShift and NVIDIA MIG

338 · Red Hat · Oct. 14, 2025, 7:37 a.m.
Summary
This blog post explores the dynamic GPU slicing capabilities provided by NVIDIA's Multi-Instance GPU (MIG) technology, integrated with Red Hat OpenShift. It guides developers through the concept of splitting a single GPU's resources into manageable slices that can be allocated to various workloads dynamically. The article features three practical demos, showcasing how to efficiently utilize GPU resources for running multiple models concurrently, with a focus on maintaining predictable performance and isolation between workloads. The content emphasizes ease of deployment in a Kubernetes environment, aiming to enhance both throughput and operational simplicity.