Run your AI cloud like a hyperscaler.

One platform to turn raw GPUs into every kind of cluster, from bare metal up to Kubernetes, Slurm, inference, and more.

Every cluster is a managed product you can ship.

Sell to customers or serve internal teams, all from one platform.

Clusters

Ship any cluster as a product

Kubernetes Clusters

Nested Clusters

Oss

Slurm Clusters

Beta

Run:AI Clusters

Ray Clusters

Inference Clusters

Dynamo · llm-d · + more

Soon

Agent Sandbox Clusters

Soon

+ more clusters

Platform for Tenant Management

Operate every tenant at scale

Cluster Templates

Capacity Management

Observability

Node Autohealing

Soon

Billing

Alpha
One stable API to consume
Machines

Provision hardware like a cloud

Bare Metal Machines

Virtual Machines

Node provisioning & lifecycle management
Drivers

Future-proof your node provisioning

Bare Metal Provisioning Drivers

+ more

Network Automation

VM Provisioning Drivers

+ more
YOUR GPU INFRASTRUCTURE
Bare metal GPU & CPU servers · Networking · Storage

Why The AI Industry’s Best Leaders Rely on vCluster

Turnkey today, APIs when you’re ready to build your own.

It’s a two-way door: start turnkey, evolve to building blocks, with no re-platforming.

Option A: Turnkey
Give tenants our UI, out of the box

Tenants self-serve through our UI and CLI.

Option B: Building blocks
Build your own experience

Call our APIs and put your own portal and brand in front.

POST /nodes
POST /tenants
POST /clusters
GET /metrics
Your portal, our APIs

Every tenant gets their own isolated cluster. Nothing else.

Standard Kubernetes
Tenants see the platform’s internals and every other tenant.
With vCluster
Each tenant sees only their own isolated cluster.
Standard Slurm
Tenants see the platform’s internals and every other tenant.
With vCluster
Each tenant sees only their own isolated cluster.
Standard Run:AI
Tenants see the platform’s internals and every other tenant.
With vCluster
Each tenant sees only their own isolated cluster.
Standard Ray
Tenants see the platform’s internals and every other tenant.
With vCluster
Each tenant sees only their own isolated cluster.

One platform. Every kind of builder.

Unlock more revenue with a hyperscaler-like experience for your customers
  • Managed clusters of every kind, from Kubernetes and Slurm to Ray and inference
  • Automate tenant, cluster, and bare metal provisioning end to end
  • Deep integration into your network and hardware layer
Launch managed clusters fast
Trusted by the fastest-growing AI cloud providers
1 min

To spin up isolated K8s environments

<45

Days from decision to production launch

100%

Data residency enforced at the infra layer

170+

Tenant clusters currently in production

Give every AI team the GPU access they need, without multiplying your infrastructure
  • Isolated clusters per team, project, or training run
  • Cloud-like self-service for ML and research teams, with no new infrastructure overhead
  • Maximize utilization with automatic node allocation based on capacity needs
Operate your GPUs like a hyperscaler
Reference Architecture: vCluster on NVIDIA DGX

“With vCluster on DGX systems, you can bring the elasticity, automation, and multi-tenancy of Kubernetes onto your on-prem infrastructure. Get the experience of the public cloud on your DGX systems.”

Provision clusters across any public or private cloud, and shift workloads freely
  • One platform to manage tenants across EKS, GKE, AKS, bare metal, and private cloud
  • Consistent developer experience regardless of the underlying infrastructure
  • Shift workloads between environments without re-architecting
Trusted by the teams at
Why vCluster

You think you can DIY this?

This isn’t a side project. Behind every deployment is 5+ years of deep infrastructure engineering, security hardening, and battle-tested operations at massive scale.

100K+
GPUs powered
1M+
CPUs powered
50+
GPU clouds & Fortune 500s
40+
Hardcore infra engineers
Resilient by design

Auto-healing control planes that keep tenants online.

Hardened and pentested to the kernel

Isolation, RBAC, and kernel-level security audits.

Real 24/7 expert support

Infra engineers on call, not a tier-1 help desk.

Always shipping

New chip and feature support continuously, not quarterly.

The best in the industry trust vCluster.

Architecting Production-Grade NVLinked GPU Clusters for AI
NVIDIA GTC 2026
Architecting Production-Grade NVLinked GPU Clusters for AI

From R&D to production: how to build a secure, multi-tenant NVLink cluster that scales.

vCluster on NVIDIA DGX Systems Reference Architecture
Ebook
vCluster on NVIDIA DGX Systems Reference Architecture

A blueprint for bringing cloud-grade elasticity and automation to NVIDIA DGX systems.

Kubernetes at Enterprise Scale: JPMorganChase, NVIDIA & vCluster on AI Infrastructure
Conference Talk
Kubernetes at Enterprise Scale: JPMorganChase, NVIDIA & vCluster on AI Infrastructure

A conversation on enterprise Kubernetes operations at scale with two of the world’s largest infrastructure teams.

How Nscale Builds Kubernetes Platforms on Bare Metal
YouTube
How Nscale Builds Kubernetes Platforms on Bare Metal

NScale on building a production GPU cloud using vCluster for tenant isolation at scale.

Featured in leading Kubernetes and platform engineering books
4 books with a chapter on vCluster
Featured in leading Kubernetes and platform engineering books

Your playbook for scaling Kubernetes securely with multi-tenancy that actually works.

Kubernetes Multi-Tenancy at Adobe Scale
YouTube
Kubernetes Multi-Tenancy at Adobe Scale

How Adobe uses vCluster to deliver isolated Kubernetes environments to internal teams.

GET STARTED

Deploy vCluster in minutes.

With a few simple commands, you can create your first cluster and define how workloads are isolated — all with a lightweight vcluster.yaml config.
Tenant clusters run on fully separate nodes with their own CNI, CSI, and control
# vcluster.yaml
privateNodes:
  enabled: true
controlPlane:
  service:
    spec:
      # could also be LoadBalancer if available
      type: NodePort
Tenant clusters share the host’s nodes and plugins
# vcluster.yaml
sync:
  fromHost:
    nodes:
      # set to true for real node specs
      enabled: false

Deploy on...

See vCluster in action.

Talk to an AI infra expert and get a live walkthrough built around your use case.