AI Workload Placement
The decision about which system, operating environment, and partition executes an AI model or agent, separately from the IBM i application supplying its data.
For an IBM i shop, workload placement distinguishes the production IBM i data source from the runtime that executes inference. The runtime can be an external service, a supported Power CPU environment, or a supported Linux and Red Hat configuration with an accelerator. The choice depends on measured latency, data controls, model requirements, and Power headroom, not the presence of IBM i alone.
Related Terms
AI Inference Partition
A separate supported operating-system partition used to execute an AI model while an IBM i partition may remain the production application and data source.
Matrix Math Acceleration (MMA)
An on-chip acceleration feature in IBM Power10 and Power11 processors that speeds up the matrix math operations used in AI inferencing, without requiring a separate GPU.
IBM i
IBM's operating system for Power servers running traditional AS/400-lineage business applications, with Db2 for i built directly into the OS.