Multi-Tenancy & Governance

Multi-tenancy (GPU)

Multi-tenancy is the ability to safely run multiple customers' or teams' workloads on shared GPU infrastructure, with strict isolation of compute, memory, and data between them.

True multi-tenancy, not just VM-level separation, is one of the harder problems in GPU infrastructure, since GPUs weren't originally designed to be shared.

How hosted·ai approaches this

hosted·ai built its scheduling model specifically to solve multi-tenant GPU workloads. See how hosted·ai schedules multi-tenant GPUaaS workloads for the technical approach.

Related terms