GPU Allocation
GPU Allocation
Stay updated with our expert articles and insights on cloud-native and AI infrastructure management and orchestration topics.
Compute Domains: Bringing Multi-Node NVLink Awareness to Kubernetes
Learn how compute domains and multi-node NVLink enable high-performance, distributed GPU workloads in Kubernetes, improving scalability, resource utilization, and AI infrastructure efficiency. Read Now
Bare Metal Isn’t a Business Model: How Cloud Providers Monetize AI Infrastructure
Dynamic Resource Allocation for GPU Allocation on Rafay's MKS (Kubernetes 1.34)
Deploy Workload using DRA ResourceClaim in Kubernetes
In this second blog, we installed a Kubernetes v1.34 cluster and deployed an example DRA driver on it with "simulated GPUs". In this blog, we’ll deploy a few workloads on the DRA enabled Kubernetes cluster to understand how "Resource Claim" and "ResourceClaimTemplates" work.
Read Now
Introduction to Dynamic Resource Allocation (DRA) in Kubernetes
In this post, we’ll look at how a new GA feature in Kubernetes v1.34 — Dynamic Resource Allocation (DRA) — aims to solve these problems and transform GPU scheduling in Kubernetes.
Read Now
Rethinking GPU Allocation in Kubernetes
Kubernetes has cemented its position as the de-facto standard for orchestrating containerized workloads in the enterprise.
Read Now