Understanding user personas

The personas of IBM Fusion Content-Aware Storage (CAS) service include Scale Admin, Data Engineer, and Developer. The Developer connects RAG application for working with the ingested content.

Scale Admins

Scale Admins are individuals responsible for managing storage software, such as IBM Storage Scale clusters and storage infrastructure. The Scale Admin deploys CAS with IBM Fusion and AI pipeline onto OpenShift® Container Platform cluster. The Scale Admin can see events for CAS in the same interface as IBM Fusion for a consistent serviceability experience​.

The key responsibilities of a Scale Admin are as follows:
  • Deploy, configure, and manage IBM Storage Scale clusters to ensure optimal performance and reliability.
  • Configure and fine-tune IBM Storage Scale to efficiently handle high-throughput, low-latency AI data processing, ensuring integration with GPUs and AI frameworks.
  • Work with Data Engineer to oversee data ingestion, preprocessing, and movement across storage tiers.
  • Monitor system health of the Scale environment, implement caching strategies, and optimize data placement to prevent bottlenecks in AI model training and deployment.
  • Leverage automation tools and scripts to streamline data lifecycle management, scaling storage resources dynamically to meet AI workload demands.

Data Engineer

Data engineers leverage data lakehouses as a flexible and scalable solution to store, process, and analyze diverse data types. Data Engineer defines a fileset on its storage content and connects it to an AI pipeline. In addition, the Data Engineer monitors the execution of the AI pipeline to understand the amount and timeliness of processing changes in content. As a Data Engineer, I can configure an Nvidia NIM pipeline to my Scale Fileset through CRs​​.

​The key responsibilities of a Data Engineer are as follows:
  • Storage of raw data in its original format. Including structured, semi-structured and unstructured data​.
  • Design and implement data ingestion pipelines that extract data from various sources, ensuring that all relevant data is available via a single interface for analysis and reporting.​
  • Maintain/curate metadata catalogs that capture information about data sources, data transformations, and data dependencies​.
  • Transform, augment, cleanse, and derive new datasets from existing data in accordance with their consumer’s needs.​
  • Define access controls, data security measures, and data privacy policies to ensure compliance with policy and regulatory requirements.
  • Enable data scientists, analysts, and other stakeholders to explore and analyze data​.

Application Developer

The Application Developer queries and chats using the client enterprise application.