SFTP Setup Guide

Follow our setup guide to sync files from your SSH server to your destination using SFTP.

Prerequisites

To set up a Fivetran SFTP connection, you need the following:

An account on an SSH server containing files with supported filetypes and encodings.
The ability to log in to this account using either a password or a key pair.

Setup instructions

Begin Fivetran configuration

In the connection setup form, select the sync strategy: Magic Folder or Merge Mode.
Enter the Destination schema name of your choice.
If you selected Merge Mode as your sync strategy, enter the Table group name. We combine this with the destination schema to form the Fivetran connection name <destination_schema>.<table_group_name>. This enables you to create multiple Merge Mode connections per destination schema. The Table group name value is used only in Fivetran and does not appear in your destination.
In the Destination schema names field, choose the naming convention you want Fivetran to use for the schemas, tables, and columns in your destination:
- Fivetran naming: Standardizes the schema, table, and column names in your destination according to the Fivetran naming conventions.
- Source naming: Preserves the original schema, table, and column names from the source system in your destination.
If you want to modify your selection, make sure you do it before you start the initial sync.

Connect

(Not applicable to Hybrid Deployment) In the Connection Method drop-down menu, you can choose to Connect directly or Connect via SSH Tunnel:
- Connect directly: Fivetran will connect directly to your SFTP Server. This is the simplest method. To connect directly, Fivetran's IP addresses should be safelisted in your firewall.
- Connect via SSH Tunnel: Fivetran will connect to a separate server in your network which provides an SSH tunnel to your SFTP Server. You must choose this option if your SFTP Server is in an inaccessible subnet.
For connections configured for Hybrid Deployment, Connect directly is pre-selected in the Connection Method drop-down menu.
(Not applicable to Hybrid Deployment) If you use a keypair for logging into your SFTP server, set the Login with keypair? toggle to ON. Make a note of the SFTP Server Public Key and proceed to the next section.
If you select Connect via SSH Tunnel, make a note of the automatically generated SSH Tunnel Public Key.
You can use either of the automatically generated public keys, SSH Tunnel Public Key or SFTP Server Public Key, for configuration as they are the same.
If you use password for logging into your SFTP server and chose Connection Method as Connect directly, proceed to the Add login details section.
(Hybrid Deployment only) If your destination is configured for Hybrid Deployment, the Hybrid Deployment Agent associated with your destination is pre-selected for the connection. To assign a different agent, click Replace agent, select the agent you want to use, and click Use Agent.

Configure your keypair in a text editor

If you want to connect via SSH Tunnel, this step is mandatory as we only support keypair based login to Tunnel Host. If you want to connect directly to your SFTP server, then this step is only required if your SSH user account uses key pairs to log in.

You need to create a group and SSH user, add the SSH user to the group, create the .ssh directory and authorized_keys file, and grant them permissions, unless they already exist and have the permissions granted. To do it, log in to your server and run the following commands:

Create group fivetran:
```
sudo groupadd fivetran
```
Create an SSH user fivetran:
```
sudo useradd -m -g fivetran fivetran
```
Switch to your fivetran user:
```
sudo su - fivetran
```
Create the .ssh directory:
```
mkdir ~/.ssh
```
Grant the .ssh directory permissions:
```
chmod 700 ~/.ssh
```
Switch to the .ssh directory:
```
cd ~/.ssh
```
Create the authorized_keys file:
```
touch authorized_keys
```
Grant the authorized_file permissions:
```
chmod 600 authorized_keys
```

Using your favorite text editor, add the SSH Tunnel public key from the Fivetran setup form to the authorized_keys file. The key must be all on one line. Make sure no line breaks are introduced when cutting and pasting.

Add the Fivetran public key to the /.ssh/Authorized_keys file on any SSH account you wish to use.

If you choose both login with keypair and connect via SSH Tunnel, make sure you configure the SFTP Server Public Key and SSH Tunnel Public Key in their respective servers.

Add login details

In the connection setup form, enter the following details:
- SFTP Server Host Address
- SFTP Server Port
- SFTP Server Username
- SFTP Server Password (not required for keypair login)
If you selected Connect via an SSH tunnel as the Connection Method, enter the following details:
- SSH Tunnel Host Address
- SSH Tunnel Port
- SSH Tunnel Username
If you entered DNS instead of IP Address in Host Address of SFTP Server, you must have ip address to DNS mapping for SFTP Server in /etc/hosts file of tunnel Host Or the name should resolve to ip using Internal DNS server.
(Optional) Base folder path - Choose the lowest common folder in a folder hierarchy that includes all the files you want to sync and enter it in the Base folder path field. This defines a specific location where Fivetran scans for files and helps ensure optimal performance. For example, if the files are in files/exports/customers/data_20251016.csv and files/exports/products/data_20251016.csv, set files/exports as the base folder path.
(Optional) Click Run connection test to validate the login credentials, the connection to the SFTP server, and the permissions on the Base folder path.
You can skip this intermediate test and proceed to the next step. However, if you choose to skip, we will perform this test once you have finished your configuration.

Add SFTP configuration

Magic Folder Mode

(Optional) In the setup form, enter your Folder Path from your SFTP server to specify the section of the file system in which you'd like Fivetran to look for files. If you don't provide a prefix, we'll search the root folder for files we can sync.
If you want to sync files from nested folders within the specified folder, set the Include subfolders toggle to ON.

Merge Mode

In the setup form, choose your configuration options. Using these configuration options, you can select subsets of your folders, specific types of files, and more to sync only the files you need in your destination. In addition, setting up multiple connections targeted at the same file system but with different options allows you to slice and dice a file system any way you'd like.

Configure files

File Handling - We process and sync all files based on the file handling option you select:
- Extract structured data into destination tables - Parse supported file types and sync structured data into destination tables. We recommend this option for most use cases.
- Replicate unstructured files Beta - Copy unstructured files in their original format without extracting data. This option is ideal for PDF documents, images, and other non-tabular file formats. Learn more about unstructured file replication in our documentation.
File Mapping - You can map the files to a destination using the following options:
Define per table
- Select Define per table.
- Click + Add files to specify destination tables and their corresponding file name pattern.
- Table name - Use names that are unique across all SFTP connections within the same destination schema.
- (Optional) File pattern - Use a regular expression as the file pattern to determine whether to sync specific files. The pattern you specify applies to everything under the prefix (base folder path). If you want to sync everything under the prefix, leave this field blank.
  For example, if under the prefix you have a folder data, which has sub-folders, subFolder1, subFolder2, etc. These sub-folders have JSON files with the format report_03/12/2050.json. Use the following regex patterns to decide whether or not to sync specific files:
  - data/.* matches all files in the data folder, including those in subfolders.
  - data/.*json matches all JSON files in the data folder, including those in subfolders.
  - data/subFolder2/report_.*\.json matches all the JSON files in the subFolder2 folder that have a name that starts with the prefix report_.. For example, report_file.json.
  - report_\d{2}/\d{2}/\d{4}\.json matches all the JSON files that begin with the prefix report_ and are followed by a date format of DD/MM/YYYY or MM/DD/YYYY. For example, report_03/12/2050.json.
    We recommend that you test your regex.
- (Optional) Archive file pattern - Use a regular expression to filter and sync files from archived folders. We sync the files in compressed archives with filenames matching the specified pattern. For example, if you specify the archive folder pattern as .*json, we will sync only the files that end in a .json file extension from the archive folder.
  You need to configure archive patterns per table. This is useful when an archive folder contains files following different naming patterns, allowing you to route each type to a specific destination table based on its pattern.
  For example, if the archive folder contains test12.json and check12.json, you can configure test.*\.json as archive pattern for Table1 to sync only test12.json to Table1, and check.*\.json for Table2 to sync only check123.json to Table2.
- (Optional) Click Preview Files to validate the file pattern and archive file pattern.
  You can skip this intermediate test and proceed to the next step. However, if you choose to skip, we will perform this test once you have finished your configuration.
- Click Save.
Dynamically extract tables
- Select Dynamically extract tables.
- Use this option to dynamically extract table names from file paths using a regular expression with a named capture group.
- Table extraction pattern - Specify a regular expression with a named capture group (?<table>...) to extract the table name from matching file paths.
  For example, if your files follow a naming pattern like 20250101/report/customers.csv, 20250101/report/orders.csv, etc., you can use the pattern \d{8}/report/(?<table>\w+)\.csv. Fivetran will automatically create separate destination tables for each unique table name extracted from the pattern (for example, customers, orders). To learn more about Dynamic File Mapping, see How to use Dynamic File Mapping?
  We recommend that you test your regex to ensure it correctly captures the table name.
- (Optional) Click Preview to validate the regex pattern and see which table names will be extracted from your files. The preview displays one matched file per table with the corresponding table name extracted from the file path.
  The preview displays the table names extracted from your files. These names will be converted according to Fivetran's naming conventions when synced to your destination. For more information, see our naming conventions documentation.
- Any new tables observed post-setup (i.e. previously unseen table values that match your pattern) will be added automatically. You can control this behavior using Schema change settings.

Dynamically extract tables mode does not support the following features: archive patterns and custom per-table file patterns.

Format

File Type - We process all files as the selected file type. Use the File pattern field (in Define per table mode) or Table extraction pattern (in Dynamically extract tables mode) to select the file extensions you want to sync. If your file type is XML, we load your XML data into the _data column without flattening it.
If your file type is CSV or TSV then enter the following details:
- (Optional) Delimiter - Specify the delimiter used in your CSV file. If your CSV file uses a custom delimiter, replace the default comma , with your specific delimiter. For example, if your file is tab-delimited, enter \t, or if it's pipe-delimited, enter |. If you leave this field blank, we'll attempt to detect the delimiter for each file automatically. However, note that automatic detection may not work in all cases. If your files sync with an incorrect number of columns or use a unique delimiter, consider specifying the delimiter. You can store files with different delimiters in the same folder. For more details on how delimiter inference works, see our documentation.
- Quote character - Typically, CSVs use double quotes " to enclose a value. Set the toggle to off if you don't want to use an enclosing character.
- Non-Standard escape character - Set the toggle to ON if your CSV generator uses non-standard ways of escaping characters like newline, delimiter, etc. Not standard in CSVs.
- Null Sequence - Set the toggle to ON if your CSVs use a special value indicating null. Specify the value indicating null only if you are sure your CSVs have a null sequence. Typically, CSVs have no native notion of a null character. However, some CSV generators have created one, using characters such as \N to represent null.
- Skip Header Lines - Use this option to skip over a fixed number of header lines at the beginning of your CSV files. Set the toggle to ON, and then in the Number of skipped header lines field, specify the number of header lines you want to skip.
- Skip Footer Lines - Use this option to skip over a fixed number of footer lines at the end of your CSV files. Set the toggle to ON, and then in the Number of skipped footer lines field, specify the number of footer lines you want to skip.
- Headerless files - Set the toggle to ON if your CSV-generating software doesn't provide a header line. Fivetran can generate generic column names and sync data rows with them.
- Line Separator - Line separators are used in CSV files to separate one row from the next. By default, we use the new line character \n as the line separator. If you use a different line separator for your CSV files, replace \n with your custom line separator.
If your file type is JSON or JSONL, then select the following:
JSON Delivery Mode - Use this option to choose how Fivetran should handle your JSON data.
- If you select Packed, we load all your JSON data into the _data column without flattening it.
- If you select Unpacked, we flatten one level of columns and infer their data types.
If your file type is XLS/XLSX/XLSM, then enter the following details:
By default, we analyze your spreadsheet to identify the cell reference. You can also opt to enter a cell reference of your choice by using the Manually provide cell reference toggle. We use the cell reference to sync all contiguous data starting from the top-left cell in all the spreadsheets matching the name.
- Analyze sheet: Identify the sample file you would want to sync. We analyze and identify the eligible data sets. To determine the cell reference correctly, do the following:
  - In the Spreadsheet to find data to be synced field, enter the path from the root folder of one of your Excel files.
  - Click Analyze sheet.
  - In the Cell reference for syncs drop-down menu, select the cell reference.
- Set the Manually provide cell reference toggle to ON to enter the cell reference.
  - Manual Cell Reference: Enter the cell reference in the '<sheetName>'!<startColumnName><startRowName> format. For example, if you want to sync data starting from cell 'C3' of the 'Data2' worksheet, enter 'Data2'!C3.
Learn more about syncing Excel files.
Primary Key used for file process and load - Use this option to let Fivetran know how you'd like to update the files in your destination. When you modify a previously synced file, the option you select determines if we should replace the rows in the destination table or append new rows to the table:
- If you select Upsert file using file name and line number, we will upsert your data using the surrogate primary keys _file and _line. If a file has a unique name, we will sync the data for that file as new data.
- If you select Append file using file modified time, we will upsert your files using surrogate primary keys _file, _line, and _modified. You can track the full history of a file or set of files and your files will have a combination of old and new data or data that is updated periodically.
- If you select Upsert file using custom primary key, you can keep the most recent version of every record, and your files will have a combination of the old and new data or data that is updated periodically. You can choose the primary keys you want to use after you save and test.
  You can't modify your primary key option once the initial sync is successful. However, if you selected Upsert file using custom primary key, you can change the columns selected as primary keys after the initial sync.

Additional options

Compression - If your files are compressed but do not have extensions indicating the compression method, you can decompress them according to the selected compression algorithm. If all of your compressed files are correctly marked with a matching compression extension (.bz2, .gz, .gzip, .tar, or .zip), you can select infer. If you select uncompressed, we do not decompress the files and sync the uncompressed files. If you choose a compression format, we decompress every file using the format you select. For example, if you have an automated CSV output system that GZIPs files to save space but saves them without a .gzip extension, you can set this field to GZIP. We will decompress every file that we examine using GZIP.
Error Handling - Use the error handling option to choose how to handle errors in your files. If you know that your files contain some errors, you can choose to skip poorly formatted lines.
- If you select skip, we ignore improperly formatted data within a file, allowing you to sync only valid data.
- If you select fail, we fail the sync with an error on finding any improperly formatted data.
  We recommend that you select fail unless you are sure that you have undesirable, malformed data.
You will receive a notification on your Fivetran dashboard if we encounter errors.
(Optional) PGP Encryption Options - Use this option to sync PGP encrypted files. Set the toggle to ON and specify the following:
- PGP Private Key - Upload the PGP secret key as an attachment.
- (Optional) Passphrase - Enter the passphrase you used to generate the key.
- (Optional) Signer's Public Key - Upload the signer's public key as an attachment. This key is used for verifying the files.
  For PGP decryption processes, we strictly comply with the RFC4880 standard. We support syncing only base64 encoded files.
  To support PGP encryption on compressed files, the file name must contain both a valid compression extension and the .pgp encryption extension. For example: sample.csv.zip.pgp — where .zip is the compression extension and .pgp is the encryption extension.
(Hybrid Deployment only) If your destination is configured for Hybrid Deployment, the Hybrid Deployment Agent associated with your destination is pre-selected for the connection. To assign a different agent, click Replace agent, select the agent you want to use, and click Use Agent.

Finish Fivetran configuration

Click Save & Test. Fivetran will take it from here and sync data from your SFTP server.
Click Confirm to ensure that the server SSH key is trusted when the initial tests run.

Fivetran tests and validates the SFTP connection. On successful completion of the setup tests, you can sync your SFTP data to your destination.

Setup tests

Depending on your sync strategy and file mapping method, Fivetran performs the following SFTP connection tests:

(Both Magic Folder Mode and Merge Mode) The Validating Connection Parameters test validates the username, password, host address, and port number you specified in the setup form.
(Both Magic Folder Mode and Merge Mode) The Connecting to SFTP Server test validates the server credentials you specified in the setup form and checks the accessibility of your SFTP server.
(Merge Mode - Define per table) The Finding tables test validates if you have specified at least one table in the files field to set up the connection.
(Merge Mode - Define per table) The Validating Regex File Pattern test validates all the file pattern regex you specified in the setup form.
(Merge Mode - Dynamically extract tables) The Validating Table extraction pattern test validates the table extraction pattern regex and ensures it includes a valid (?<table>...) named capture group.
(Merge Mode - Define per table) The Validating Archive Pattern test validates the archive pattern regex you specified in the setup form. We perform this test only if you specify a regex in the Archive File Pattern field.
(Merge Mode) The Validating EscapeChar test validates the escape character you specified for your CSV files and checks the length of the character, which must not be more than one. We perform this test only if you specify an escape character in the Escape Character field.
The Validating Infer FileType test validates whether infer is added as a value in the file_type parameter for connections created using the API. We perform this test only if you have set up your connection using the API.
(Merge Mode) The PGP Support test validates whether the connection can successfully retrieve a minimum of one sample file and a maximum of ten sample files from FTP and decrypt them using the PGP keys you uploaded. We perform this test only if you set the PGP Encryption Options toggle to ON.
(Merge Mode) The Multi-Character Delimiter Support test validates the length of the delimiter, which must be within 15 characters. We perform this test only if you specify the delimiter for your CSV files in the Delimiter field.
(Merge Mode - Define per table) The Finding Matching Files test checks if the connection can successfully retrieve a minimum of one sample file and a maximum of five sample files for each of the tables you specified in the setup form.
(Merge Mode - Dynamically extract tables) The Finding Matching Files test checks if the connection can successfully discover tables and retrieve a minimum of one sample file and a maximum of three sample files for each discovered table (up to 5 tables).