Azure Blob Storage

Polytomic connects Azure Blob Storage to your data warehouse, databases, and SaaS tools for file based ETL and reverse ETL. Read CSV, JSON, and Parquet files into your data warehouse and other systems, and write query results back to Blob Storage as scheduled files, without writing code.

Azure Blob Storage logo

CSV, JSON, and Parquet

Read and write all three, with gzip, zstd, and bzip2 on the way in.

Recursive path patterns

Match nested folders and turn parts of the path into real columns.

Flexible authentication

Account key, SAS token, service principal, or delegated user OAuth.

Example workflows

Load Blob Storage files into your data warehouse and other systems

What you can sync

Move data between Azure Blob Storage and your data warehouse, databases, and SaaS tools, reading files in and writing scheduled output back out.

From Azure Blob Storage

CSV, JSON arrays, JSON lines, and Parquet
Gzip, zstd, and bzip2 compressed CSV and JSON
Recursive glob patterns with path values captured as columns
One file per table, many files as one table, or discovered tables
Headerless CSV and configurable skipped lines

To Azure Blob Storage

CSV, JSON lines, JSON documents, and Parquet output
Replicate mode replacing a single stable blob
Snapshot mode writing timestamped blobs
Incremental append writing changed records as new blobs
Container prefix and configurable output subfolder

How it works

Point Polytomic at a container and it discovers the files there, infers schemas, and loads them into your data warehouse, databases, or SaaS tools. In the other direction it writes query results back to Blob Storage on a schedule, in whichever format the consuming system expects.

Path patterns are what make recurring drops workable. Patterns can recurse through nested folders, gather many files into one table, and capture parts of the path, such as a date or a region, into real columns rather than context that disappears on load.

  • Connect with an account key, SAS token, service principal, or OAuth
  • Read CSV, JSON, JSON lines, and Parquet, including compressed files
  • Match files with recursive glob and capture patterns
  • Group files as individual tables, one combined table, or discovered tables
  • Write output as replicate, snapshot, or incremental append
  • Monitor sync health, record volume, and failures from one interface

How to set up

Polytomic supports four authentication paths: a storage account key, a shared access signature entered in the access key field, service principal credentials using tenant, client ID and secret, or delegated user OAuth. The connection can be scoped to a container and prefix.

How to get connected
  1. 1Decide which authentication method suits your environment
  2. 2In Polytomic, navigate to Connections, then Add Connection, then Azure Blob Storage
  3. 3Enter your storage account and credentials
  4. 4Set the container and any prefix that scopes the connection
  5. 5Add a glob or capture pattern if you want to match a set of files
  6. 6Test the connection and click Save

See the documentation for full setup details.

Frequently asked questions

Why teams choose Polytomic

No engineering required

Set up and manage syncs without writing code.

Flexible data modeling

Use SQL to define exactly what data gets synced.

Handles scale automatically

Supports large datasets with incremental syncs and bulk APIs.

Start syncing your data with Azure Blob Storage in minutes

No credit card required. Free trial available.

More integrations beyond Azure Blob Storage