Introduction
The Sizing Guide for SAP Data Intelligence Cloud explains the sizing for the diagnostics, cluster applications, tenant applications, user applications, and pipeline components of SAP Data Intelligence Cloud.
SAP Data Intelligence Cloud is a comprehensive data management solution that connects, discovers, and enriches disjointed data assets and then transforms them into actionable business insights at enterprise scale.
SAP Data Intelligence Cloud enables the creation of data warehouses from diverse enterprise data, simplifies the management of IoT (Internet of Things) data streams, and facilitates scalable machine learning. SAP Data Intelligence allows you to leverage your business applications to become an intelligent enterprise and provides a holistic, unified way to manage, integrate, and process all of your enterprise data.
Functions of SAP Data Intelligence Cloud
-
Data Governance:
Discover information by indexing, publishing, and profiling your datasets. Extract metadata and lineage to learn about the source and transformations of the dataset. Organize and label your datasets to help you find data quickly. Use data preparation to find data quality issues, correct and standardize data, and then output the data for analysis.
-
Data Pipelines:
The SAP Data Intelligence Modeler helps create data processing pipelines (graphs) and provides a runtime and design-time environment for data-driven scenarios. The Modeler reuses existing coding and libraries to orchestrate data processing in distributed landscapes.
-
Replications:
Replicate datasets from a source system to a target system.
SAP Data Intelligence Cloud Components
Multiple components comprise SAP Data Intelligence Cloud and factor into sizing.
The following table lists the system components that constitute the system management and central persistencies of SAP Data Intelligence Cloud, including an SAP HANA database, a storage gateway (to abstract connectivity to storages) and a user account and authentication server (UAA).
|
System Component |
Description |
|---|---|
|
Diagnostics |
Services for monitoring of performance metrics and log messages. |
|
Cluster applications |
Applications running one instance per cluster. This includes the Central Connection Management (CCM) and the Replication Management Service (RMS). |
|
Tenant applications |
SAP Data Intelligence Applications running one instance per tenant. This includes, for example, the Launchpad and the Metadata Explorer. |
|
User applications |
SAP Data Intelligence Applications running one instance per user. This includes the Pipeline Modeler. |
|
Pipelines |
All workloads, including those under the category of Data Governance, are executed in terms of pipelines. Pipelines run in separate Kubernetes pods (at least one pod per pipeline). |
The following diagram shows the structure of SAP Data Intelligence: