Table of Contents


  1. Preface
  2. Introduction to Informatica Big Data Management
  3. Mappings in the Hadoop Environment
  4. Mapping Sources in the Hadoop Environment
  5. Mapping Targets in the Hadoop Environment
  6. Mapping Transformations in the Hadoop Environment
  7. Processing Hierarchical Data on the Spark Engine
  8. Configuring Transformations to Process Hierarchical Data
  9. Processing Unstructured and Semi-structured Data with an Intelligent Structure Model
  10. Stateful Computing on the Spark Engine
  11. Monitoring Mappings in the Hadoop Environment
  12. Mappings in the Native Environment
  13. Profiles
  14. Native Environment Optimization
  15. Cluster Workflows
  16. Connections
  17. Data Type Reference
  18. Function Reference
  19. Parameter Reference

Cluster Workflows Process

Cluster Workflows Process

Creation of a cluster workflow requires administrator and developer tasks.
The following image shows the process to create, configure, and run a cluster workflow:
First, an administrator completes the following steps:
  1. Verify domain and cloud platform prerequisites.
  2. Create the cluster provisioning configuration on the domain.
  3. Create a Hadoop connection for the cluster workflow to use.
Then, a developer completes the following steps:
  1. Create the workflow and the Create Cluster task.
  2. Create Mapping tasks and associate them with mappings that you prepared. Optionally add Command tasks and other tasks to the workflow.
  3. Optionally add a Delete Cluster task to the workflow.
  4. Deploy and run the workflow.

Updated October 23, 2019