Skip to main content
The Polystack On-Prem AI Infrastructure Platform delivers a complete, private AI stack on your own hardware. GPU-capable virtual machines and Kubernetes clusters, a built-in operator catalog, runtime security scanning, a ready-to-use MLOps toolchain, and accelerator-level performance tuning are managed together from one console, across NVIDIA, Intel Gaudi, and AMD accelerators.
The platform builds on Polystack Compute for GPU-capable virtual machines and on Polystack Kubernetes as a Service for GPU-enabled container clusters.

Platform Capabilities

GPU-Capable VMs and Containers

GPU passthrough and GPU sharing for virtual machines, plus Kubernetes clusters with dedicated GPU node groups.

Operator Framework

Operator Lifecycle Manager and the OperatorHub.io catalog, installed by default on every Kubernetes cluster.

Runtime Vulnerability Scanning

StackRox runtime protection across all clusters, with Harbor and built-in Trivy for image and registry scanning.

MLOps Platform

Kubeflow, KServe, and MLflow, delivered as a one-click AI cluster template.

GPU and HPU Performance Optimization

CPU pinning, NUMA affinity, hugepages, vendor GPU operators, GPUDirect RDMA, and accelerator telemetry.

Multi-Vendor Accelerator Support

NVIDIA, Intel Gaudi, and AMD accelerators, each with its own flavor, cluster template, and operator.

Unified Console

One portal and one single sign-on for CPU and GPU workloads, VMs, and Kubernetes clusters.

Platform Architecture


How the Platform Fits Together

Accelerators are exposed to workloads

Polystack Compute presents NVIDIA, Intel Gaudi, and AMD accelerators to instances through PCI passthrough, NVIDIA vGPU, or AMD SR-IOV, using a dedicated flavor per vendor.

Kubernetes clusters consume GPU flavors

Polystack K8SaaS provisions clusters with GPU node groups on those flavors, and a vendor-specific operator runs on each node group.

Every cluster ships with platform services

Operator Lifecycle Manager, the OperatorHub.io catalog, and StackRox Sensor and Collector are deployed to every cluster automatically.

AI teams start from a ready-made AI cluster

The AI cluster template adds the GPU Operator, Kubeflow, MLflow, and Harbor in a single deployment.

Everything is managed from one portal

The Polystack Dashboard and Rancher sit behind one single sign-on and are presented as a single unified portal.

Related Services

Kubernetes as a Service

Cluster templates, node groups, and lifecycle management for the Kubernetes clusters the platform runs on.

Compute

Instances, flavors, and hypervisor configuration behind GPU-capable virtual machines.

Monitoring

Prometheus and Grafana dashboards that receive accelerator telemetry.

Identity and Access

Projects, users, and roles that back the platform’s single sign-on.