NVIDIA Reference Architecture | Rafay

The Definitive PaaS Reference Architecture for GPU Cloud Providers & Enterprises

Enterprises and Cloud Providers leverage the Rafay Platform to deliver self-service consumption of AI and cloud-native infrastructure to developers and data scientists.

By delivering a GPU PaaS™ experience, teams can:

Enabling Self-Service Consumption of GPUs and AI Applications with Nvidia

Self-Service GPU & AI App Consumption with NVIDIA and Rafay | Platform Demo - YouTube

Self-Service GPU & AI App Consumption with NVIDIA and Rafay | Platform Demo

Featured Resources

Operationalizing AI Fabrics with Aviz ONES, NVIDIA Spectrum-X, and Rafay
Discover the new AI operations model available to enterprises that enables self-service consumption and cloud-native orchestration for developers.
Learn More

The Definitive GPU PaaS Reference Architecture
Understand what it takes to deliver the right GPU infrastructure to your business.
Learn More

Unlock Your AI Potential with Cisco and Rafay: Transform AI PODs into a Self-Service GPU Cloud
Cisco provides AI-optimized infrastructure. Rafay makes it usable across teams, tenants, and use cases in days.
Learn More

The CIO’s guide to scalable, compliant, and developer-ready AI deployment
Orchestrating the future of AI: The CIO’s guide to scalable, compliant, and developer-ready AI deployment
Learn More

Rafay Named Outperformer in 2025 GigaOm Radar Report for Managed Kubernetes
The latest Radar report from GigaOm, Managed Kubernetes Rafay is ranked as an “Outperformer” for its solution.
Learn More

Building AI Value within Borders
Rafay's central orchestration platform facilitates efficient, self-service infrastructure and AI application management.
Learn More

GPU cloud evaluation report
Evaluating how the Rafay Platform delivers a GPU cloud for enterprises and cloud service providers by PivotNine.
Learn More

.png)

How Enterprise Platform Teams Can Accelerate AI/ML Initiatives
This paper explores the key challenges that organizations experience supporting these initiatives, as well as best practices for successfully leveraging Kubernetes to accelerate AI/ML projects.
Learn More