2.3 Administering Run:ai
Projects, departments, roles and node pools.
Key points
A project can be a team, a person or an initiative. In Kubernetes, each project becomes a namespace.
What NVIDIA says (2)
“NVIDIA Run:ai uses Projects as the primary organization management unit.”
“Projects are manifested as Kubernetes namespaces.”
Departments sit above projects. They let you allocate quota across projects and apply policies at department level.
What NVIDIA says (1)
“Departments group multiple projects under a shared organizational scope.”
A subject is a user, group or service account. A scope is the part of the organization the role applies to. Run:ai has predefined roles and allows custom ones.
What NVIDIA says (1)
“A role defines a set of permissions that can be assigned to a subject in a scope”
A node pool groups nodes by a label, such as GPU type. With over quota enabled, projects can still use a new pool before getting quota there.
What NVIDIA says (1)
“Once created, the new node pool is automatically assigned to all projects and departments with a quota of zero GPU resources, unlimited CPU resources, and over quota enabled”
Each node pool has its own scheduler instance. Workloads sent to a pool are scheduled by that instance.
What NVIDIA says (1)
“Creating a new node pool creates a new instance of the NVIDIA Run:ai Scheduler .”
Key terms
- NVIDIA Run:ai: A Kubernetes platform that schedules AI workloads and shares GPUs between teams by quota.
- Run:ai project: The main Run:ai unit for a team: it holds a GPU quota and maps to a Kubernetes namespace.
- Run:ai department: A group of Run:ai projects under one shared scope.
- Node pool: A Run:ai set of nodes with its own scheduler instance and quotas.
Sample question
What is the primary unit Run:ai uses to give a team GPU quota, scheduling rules and access?
Show the answer
Answer: A project
A project can be a team, a person or an initiative. In Kubernetes, each project becomes a namespace.
What NVIDIA says (2)
“NVIDIA Run:ai uses Projects as the primary organization management unit.”
“Projects are manifested as Kubernetes namespaces.”
Practice 2.3 (5 questions) Full Administration guide
← 2.2 AI data center architecture · 2.4 Administering Kubernetes for GPUs →