CSV, JSON, and Parquet
Read and write all three, with gzip, bzip2, and zstd on the way in.
Polytomic connects Google Cloud Storage to your data warehouse, databases, and SaaS tools for file based ETL and reverse ETL. Read CSV, JSON, and Parquet files into your data warehouse and other systems, and write query results back to your bucket on a schedule, without writing code.
Read and write all three, with gzip, bzip2, and zstd on the way in.
Match nested folders and capture parts of the path as real columns.
One of the most common ways teams load a Cloud Storage lake into a data warehouse.
Load Cloud Storage files into your data warehouse and other systems
Move data between Google Cloud Storage and your data warehouse, databases, and SaaS tools, reading files in and writing scheduled output back out.
Point Polytomic at a bucket and it discovers the files there, infers schemas, and loads them into your data warehouse, databases, or SaaS tools. In the other direction it writes query results back to Cloud Storage on a schedule in whichever format the consuming system expects.
Path patterns are what make recurring drops workable. Patterns can recurse through nested folders, gather many files into one table, and capture parts of the path such as a date or a region into real columns rather than context that disappears on load.
Polytomic authenticates to Cloud Storage with a service account key, and derives the project and service account identity from it. Grant that service account access to the buckets you want to read or write, and scope the connection to a bucket and optional prefix.
See the documentation for full setup details.
Set up and manage syncs without writing code.
Use SQL to define exactly what data gets synced.
Supports large datasets with incremental syncs and bulk APIs.
No credit card required. Free trial available.