GPU Allocation

GPU Allocation

Stay updated with our expert articles and insights on cloud-native and AI infrastructure management and orchestration topics.

Compute Domains: Bringing Multi-Node NVLink Awareness to Kubernetes

Learn how compute domains and multi-node NVLink enable high-performance, distributed GPU workloads in Kubernetes, improving scalability, resource utilization, and AI infrastructure efficiency. Read Now

Bare Metal Isn’t a Business Model: How Cloud Providers Monetize AI Infrastructure

Read Now

.png)](https://pr@rafay.co/ai-and-cloud-native-blog/dynamic-resource-allocation-for-gpu-allocation-on-rafays-mks-kubernetes-1-34)

Dynamic Resource Allocation for GPU Allocation on Rafay's MKS (Kubernetes 1.34)

Read Now

Deploy Workload using DRA ResourceClaim in Kubernetes

In this second blog, we installed a Kubernetes v1.34 cluster and deployed an example DRA driver on it with "simulated GPUs". In this blog, we’ll deploy a few workloads on the DRA enabled Kubernetes cluster to understand how "Resource Claim" and "ResourceClaimTemplates" work.
Read Now

Introduction to Dynamic Resource Allocation (DRA) in Kubernetes

In this post, we’ll look at how a new GA feature in Kubernetes v1.34 — Dynamic Resource Allocation (DRA) — aims to solve these problems and transform GPU scheduling in Kubernetes.
Read Now

Rethinking GPU Allocation in Kubernetes

Kubernetes has cemented its position as the de-facto standard for orchestrating containerized workloads in the enterprise.
Read Now