AI Optimizer for IBM Z and IBM LinuxONE

Designed to Scale and Optimize GenAI Inferencing

3D render of high-tech microchip featuring a layered architecture, including connectors and circuit boards

Overview

IBM AI Optimizer for IBM Z and IBM LinuxONE is the unified AI inferencing stack for the IBM Spyre Accelerator - bringing together model on-boarding, routing, and monitoring into a single, integrated solution on IBM Z and IBM LinuxONE.

Key Features

All-in-One Inferencing Stack

An integrated software appliance delivered as a single LPAR image - including the operating system, curated AI models, container runtime, observability tooling, and a management UI. This unified approach reduces operational complexity, accelerates deployment timelines, and ensures consistency across environments.

Illustration of a performance KPI speedometer

Real-Time Monitoring & Visualizations

Gain full visibility into GenAI inferencing across IBM Z with enterprise‑grade observability. Built‑in Prometheus and Grafana dashboards provide deep insights into:

  • Inferencing latency and performance
  • Hardware and Spyre utilization
  • Model usage and cross‑application activity
  • Bottlenecks and anomalies identification

This transparency helps eliminate over‑provisioning, streamline capacity planning, and drive smarter infrastructure investment.

Optimized inferencing (for models on Spyre)

AI Optimizer registers models running on Spyre for optimization. Users can configure their own routing strategies or rely on the built‑in intelligent router, which considers performance, availability, and usage patterns. Semantic tagging allows grouping of models for use‑case‑aligned routing thus providing more flexibility on inferencing requests.

Inference Router in AI Optimizer for Z - Product dashboard Screenshot

External LLM registration

Models deployed outside IBM Z or LinuxONE can be registered, tagged, grouped, and monitored along with on‑platform models. This provides a unified operational view of GenAI inferencing across hybrid environments, ensuring consistency in governance and performance tracking.

Streamlined Deployment of Generative AI for IBM Z

AI Optimizer for Z automates installation and configuration of key IBM Z Gen AI components and products, such as IBM watsonx Assistant for Z, ensuring fast and reliable setup. It validates infrastructure and provides a health dashboard for easy monitoring. This reduces complexity and accelerates time to production.

 

Illustration of two people working on laptops with code screens
Related Products IBM watsonx Assistant for Z

Simplify and transform how your users interact and manage mainframe with AI.

Learn more
AI accelerators

Accelerate AI innovation at scale with IBM infrastructure.

Learn more

Resources

Support Documentation
Community
Solution Brief
Take the next step

Discover how to use AI and machine learning to convert data from every transaction into real-time insights.