Open-source data movement for ELT pipelines and AI agents — from APIs, databases & files to warehouses, lakes, and AI applications. Both self-hosted and Cloud.
Hundreds of connectors syncing sources into warehouses, and a builder for new ones.
Hosted connectors that sync data from applications and databases into warehouses. Un produit fermé de Fivetran.
| Outil | Remplacement | Étoiles | Licence | Conditions | Auto-hébergeable | Langage | Dernière release | Dernier push |
|---|---|---|---|---|---|---|---|---|
| Airbyte | Remplacement complet | 22k | Other | Source disponible | Oui | Python | v2.0.0 | 2026-10-06 |
| DataX | Partiel | 17k | Other | Open source | Oui | Java | datax_v202309 | 2026-07-07 |
| Debezium | Partiel | 13k | Apache-2.0 | Open source | Oui | Java | v3.7.0.Final | 2026-10-06 |
| Apache SeaTunnel | Partiel | 9.7k | Apache-2.0 | Open source | Oui | Java | v3.0.0 | 2026-10-06 |
| CloudQuery | Partiel | 6.5k | MPL-2.0 | Open source | Oui | Go | cli-v6.43.0 | 2026-10-06 |
| dlt | Partiel | 5.9k | Apache-2.0 | Open source | Oui | Python | 1.30.0 | 2026-10-06 |
| ingestr | Partiel | 4k | Other | Source disponible | Oui | Go | v1.1.63 | 2026-10-06 |
| PeerDB | Partiel | 3.3k | AGPL-3.0 | Open source | Oui | Go | v0.37.11 | 2026-10-06 |
| Meltano | Partiel | 2.6k | MIT | Open source | Oui | Python | v4.4.0 | 2026-10-06 |
| OLake | Partiel | 1.5k | Apache-2.0 | Open source | Oui | Go | v0.11.3 | 2026-10-06 |
| Estuary Flow | Partiel | 982 | Other | Source disponible | Oui | Rust | v0.6.13 | 2026-10-06 |
| Sling | Partiel | 912 | GPL-3.0 | Open source | Oui | Go | v1.6.4 | 2026-10-05 |
12 alternatives
Open-source data movement for ELT pipelines and AI agents — from APIs, databases & files to warehouses, lakes, and AI applications. Both self-hosted and Cloud.
Hundreds of connectors syncing sources into warehouses, and a builder for new ones.
DataX是阿里云DataWorks数据集成的开源版本。
Batch sync between databases, warehouses and file stores through reader and writer plugins; no SaaS connectors or scheduler.
Change data capture for a variety of databases. Please log issues at https://github.com/debezium/dbz/issues.
Change data capture from databases into Kafka or other sinks; no SaaS application connectors.
SeaTunnel is a multimodal, high-performance, distributed, massive data integration tool.
Batch and streaming sync across databases, warehouses and SaaS sources, run on its own engine, Flink or Spark.
Data pipelines for cloud config and security data. Build cloud asset inventory, CSPM, FinOps, and vulnerability management solutions. Extract from AWS, Azure, GCP, and 70+ cloud and SaaS sources.
ELT from cloud provider APIs and SaaS sources into databases, run from a CLI with plugins.
data load tool (dlt) is an open source Python library that makes data loading easy 🛠️
A Python library for loading data into warehouses, with schema inference and incremental loads.
ingestr is a CLI tool to copy data between any databases with a single command seamlessly.
Copies tables between databases and SaaS sources with one command, without a server.
Fast, Simple and a cost effective tool to replicate data from Postgres to Data Warehouses, Queues and Storage
CDC replication from PostgreSQL into ClickHouse, Snowflake, BigQuery and queues; PostgreSQL sources only.
Meltano: the declarative code-first data integration engine that powers your wildest data and ML-powered product ideas. Say goodbye to writing, maintaining, and scaling your own API integrations.
Runs Singer taps and targets as code from a CLI, without a hosted UI.
OLake - Fastest Databases, Kafka & S3 Replication to Apache Iceberg with Table optimization (Called OLake Fusion). ⚡ Efficient, quick and scalable data ingestion for real-time analytics. Supported sources : Postgres, MongoDB, MySQL, Oracle, MSSql, DB2, Kafka, S3.
Replicates databases and Kafka into Apache Iceberg tables; few SaaS sources.
🌊 Continuously synchronize the systems where your data lives, to the systems where you _want_ it to live, by managing your data flows with Estuary. 🌊
Real-time CDC and batch connectors into warehouses and other destinations.
Sling is a CLI tool that extracts data from a source storage/database and loads it in a target storage/database.
Replicates between databases, files and object storage from a CLI and YAML; far fewer SaaS sources.