You can create a new pipeline that writes back data from an Anaplan Data Orchestrator dataset to Snowflake.
Prerequisites
You need a connection to Snowflake to create a pipeline. Make sure you meet the prerequisites in these sections before you create a connection to Snowflake and a writeback pipeline.
Connectivity prerequisites
Writeback pipeline prerequisites
Create a connection with Snowflake
Use the Snowflake connector in Data Orchestrator to create a connection.
You need your Snowflake credentials to connect the Snowflake data with Data Orchestrator. View the Snowflake documentation for more information about your credentials.
To create a connection:
- Select Data Orchestrator from the top-left navigation menu.
- Choose a dataspace from the list.
- Select Connections from the left-side panel.
- Select Create connection.
- Select the Snowflake connector and then select Next.
If you can't find the connector, enter a search term in the Find... field. - Enter these details on the Connection details screen, and then select Next:
- Name: Create a name for your connection. The name can contain alphanumeric characters and underscores.
- Description: Enter a description about your connection.
- Enter your Snowflake credentials on the Connection credentials screen, and then select Next.
For information about the fields on the Connection credentials screen, see Authentication options for Snowflake connections. - After the connection test is complete, select Done.
Create a writeback pipeline
When you set up the writeback pipeline, you'll use your connection to export either a source dataset or a transformation view from Data Orchestrator to Snowflake.
To create a writeback pipeline:
- Select Data Orchestrator from the top-left navigation menu.
- Choose a dataspace from the list.
- Select Pipelines from the left-side panel.
- Select Create pipeline.
- Enter a Name for your pipeline and then select Create.
You are taken to the pipeline designer view. - Select the Source icon, and then complete these steps in the right-side panel:
- Select Anaplan from the Connection type dropdown.
- Select Datasets from the Choose connection dropdown.
- Enter a new Label to change the source display name in the designer view.
- Select Source location > Source, choose a Data Orchestrator dataset, and then select Confirm.
The dataset is used as the source for your pipeline.
- Optionally, select the add icon that appears between the Source and Sink nodes.
You can add steps to your pipeline to process data. - Select the Sink icon, and then complete these steps in the right-side panel:
- Select Snowflake from the Connection type dropdown.
- Select the Snowflake connection you created from the Choose connection dropdown.
- Enter a new Label to change the sink name that displays in the designer view.
- Select Target location > Table, select a Snowflake table, and then select Done.
- Select Target mapping > Mapping, map the Data Orchestrator source dataset values to the Snowflake target values, and then select Done.
- Select a Write option for the target table: Append, Full replace, or Upsert.
This determines how data is written to the target table. See the Write options section below for more details.
- Select Publish, and then select Run to execute the data transfer.
Write options
Review this table to determine which write option to select.
| Write options | Description |
| Append | Adds all rows from the source data to the existing rows in the target table. If you select Append, a staging table in Snowflake is required:
See the Staging table section below for more information. Note: Append load isn't suitable for target tables with primary keys. |
| Full replace | Deletes all existing data in the target table and replaces it with the source data. If you select Full replace, a staging table in Snowflake is recommended and selected by default:
See the Staging table section below for more information. |
| Upsert | Updates existing rows and adds new rows based on a specified key. If you select Upsert, a staging table in Snowflake is required:
See the Staging table section below for more information. |
Staging tables
Staging tables are used to compare and process data before Data Orchestrator loads the data to the final target table in Snowflake.
If you want to use a staging table, you have the option to:
- Automatically let Data Orchestrator create a staging table in Snowflake.
- Manually specify an existing staging table in Snowflake.
This table describes how staging tables are used in Snowflake for each option.
| Staging table options | Results |
| Automatically create a staging table in Snowflake | After you run the pipeline:
|
| Manually specify an existing staging table from Snowflake | After you run the pipeline:
|
Note: To prevent data mismatch errors with Snowflake, make sure your data in Data Orchestrator and in Snowflake have a schema alignment. This means the tables in Data Orchestrator and in Snowflake must have the same columns and data types.
Verify the writeback in Snowflake
Once the pipeline is successfully completed, you can log in to your Snowflake account to verify the data load.
Navigate to the target database, schema, and table specified in the pipeline configuration. The data written from the Data Orchestrator dataset is available in the target table, based on the selected write option (append, full replace, or upsert).