- Support levelCertified
- Available onSelf-Managed Enterprise
- Connector version0.3.52
- Release stageAlpha
One connector
Everything S3 Data Lake can do in Airbyte
Load from any source
Move data into S3 Data Lake from 600+ Airbyte sources on a schedule you control.
Incremental syncs
Pull only the records that changed since the last run instead of reloading everything.
Cloud or self-hosted
Run the S3 Data Lake connector on Self-Managed Enterprise.
Specifications
What to know before you load into S3 Data Lake
Sync capabilities
- Full Refresh SyncSupported
- Incremental SyncSupported
- Sources600+ Airbyte connectors
What you'll need
- S3 Bucket NameThe name of the S3 bucket that will host the Iceberg data.
- S3 Bucket RegionThe region of the S3 bucket. See here for all region codes.
- Warehouse LocationThe root location of the data warehouse used by the Iceberg catalog. Typically includes a bucket name and path within that bucket. For AWS Glue and Nessie, must include the storage protocol (such as "s3://" for Amazon S3).
- Main Branch NameThe primary or default branch name in the catalog. Most query engines will use "main" by default. See Iceberg documentation for more information.
- Catalog TypeSpecifies the type of Iceberg catalog (e.g., NESSIE, GLUE, REST, POLARIS) and its associated configuration.
Common questions
Didn't find your answer?
Please don't hesitate to reach out.
What is ETL?
What data can you extract from S3 Data Lake?
How do I transfer data from S3 Data Lake?
What are top ETL tools to transfer data from S3 Data Lake?
What is ELT?
Difference between ETL and ELT?
Start moving S3 Data Lake data today
Free for 14 days on Airbyte Cloud. Set up the S3 Data Lake connector once and let Airbyte keep it in sync.