The platform builds on Polystack Compute for GPU-capable virtual
machines and on Polystack Kubernetes as a Service for GPU-enabled
container clusters.
Platform Capabilities
GPU-Capable VMs and Containers
GPU passthrough and GPU sharing for virtual machines, plus Kubernetes clusters with
dedicated GPU node groups.
Operator Framework
Operator Lifecycle Manager and the OperatorHub.io catalog, installed by default on every
Kubernetes cluster.
Runtime Vulnerability Scanning
StackRox runtime protection across all clusters, with Harbor and built-in Trivy for
image and registry scanning.
MLOps Platform
Kubeflow, KServe, and MLflow, delivered as a one-click AI cluster template.
GPU and HPU Performance Optimization
CPU pinning, NUMA affinity, hugepages, vendor GPU operators, GPUDirect RDMA, and
accelerator telemetry.
Multi-Vendor Accelerator Support
NVIDIA, Intel Gaudi, and AMD accelerators, each with its own flavor, cluster template,
and operator.
Unified Console
One portal and one single sign-on for CPU and GPU workloads, VMs, and Kubernetes clusters.
Platform Architecture
How the Platform Fits Together
Accelerators are exposed to workloads
Polystack Compute presents NVIDIA, Intel Gaudi, and AMD accelerators to instances through
PCI passthrough, NVIDIA vGPU, or AMD SR-IOV, using a dedicated flavor per vendor.
Kubernetes clusters consume GPU flavors
Polystack K8SaaS provisions clusters with GPU node groups on those flavors, and a
vendor-specific operator runs on each node group.
Every cluster ships with platform services
Operator Lifecycle Manager, the OperatorHub.io catalog, and StackRox Sensor and Collector
are deployed to every cluster automatically.
AI teams start from a ready-made AI cluster
The AI cluster template adds the GPU Operator, Kubeflow, MLflow, and Harbor in a single
deployment.
Everything is managed from one portal
The Polystack Dashboard and Rancher sit behind one single sign-on and are presented as a
single unified portal.
Related Services
Kubernetes as a Service
Cluster templates, node groups, and lifecycle management for the Kubernetes clusters the
platform runs on.
Compute
Instances, flavors, and hypervisor configuration behind GPU-capable virtual machines.
Monitoring
Prometheus and Grafana dashboards that receive accelerator telemetry.
Identity and Access
Projects, users, and roles that back the platform’s single sign-on.
