Before You Start

MindCluster allows you to deploy inference jobs through Infer Operator for scheduling and rescheduling of faulty instances.

This chapter only describes the principles of related features and corresponding configuration examples. You can refer to the configuration examples to deploy Infer Operator inference jobs.

Prerequisites

Before deploying an Infer Operator inference job, ensure that the related components have been installed. If they are not installed, refer to the Installation and Deployment chapter for instructions.

  • Volcano
  • Ascend Device Plugin
  • Ascend Docker Runtime
  • Infer Operator
  • ClusterD
  • NodeD (Optional)

Supported Product Forms

  • Atlas 800I A2 inference server
  • Atlas 800I A3 SuperPoD server

Usage Methods

MindCluster supports deploying Infer Operator inference jobs in the following ways.