Ingestion
Configure MotherDuck as the destination for your data in the following data ingestion tools.
For Python pipelines, dltHub's dlt library has a built-in MotherDuck destination and runs on MotherDuck compute as a Flight, so you can go from a source connector to a scheduled pipeline without standing up separate infrastructure. Managed platforms like Fivetran, Airbyte, and Estuary cover connectors you'd rather not maintain yourself.
Airbyte
Airbyte is a data integration platform for connecting data sources to warehouses. It integrates with MotherDuck for loading data from operational systems, APIs, files, or event streams.
Apache Kafka
Get Kafka topics into MotherDuck by materializing them as Iceberg tables, sinking them to object storage, or using a streaming ingestion partner.
Apache Spark
Write Spark DataFrames into a MotherDuck DuckLake database with the DuckLake Spark connector, or exchange data through Parquet files and the Postgres endpoint.
Artie
Artie is a fully managed CDC streaming platform that allows you to replicate data from your source database to your destination in real-time. It integrates with MotherDuck for loading data from operational systems, APIs, files, or event streams.
Ascend.io
Ascend.io is a data integration platform for connecting data sources to warehouses. It integrates with MotherDuck for loading data from operational systems, APIs, files, or event streams.
AWS Glue
AWS Glue is a serverless data integration service for preparing and moving data with Spark jobs, crawlers, and the AWS Glue Data Catalog. AWS Glue jobs can connect to MotherDuck through the MotherDuck Postgres endpoint using Glue's PostgreSQL JDBC support.
Bytewax
Bytewax is a stream processing platform for building and managing data pipelines. It integrates with MotherDuck for loading data from operational systems, APIs, files, or event streams.
CloudQuery
CloudQuery is a data integration platform for connecting data sources to warehouses. It integrates with MotherDuck for loading data from operational systems, APIs, files, or event streams.
dltHub (dlt)
dltHub builds dlt, the open-source Python library for data pipelines. MotherDuck is a first-class dlt destination for loading REST APIs, databases, and files.
Estuary
Real-time data integration platform for streaming data between systems. It integrates with MotherDuck for loading data from operational systems, APIs, files, or event streams.
Expanso
Data integration platform for connecting data sources to warehouses. It integrates with MotherDuck for loading data from operational systems, APIs, files, or event streams.
Fivetran
Automated data integration platform for connecting data sources to warehouses. It integrates with MotherDuck for loading data from operational systems, APIs, files, or event streams.
Hevo
Hevo is a data integration platform for connecting data sources to warehouses. It integrates with MotherDuck for loading data from operational systems, APIs, files, or event streams.
HubSpot
Load HubSpot contacts, companies, deals, and pipelines into MotherDuck on a schedule with a Flight that runs dlt's HubSpot source.
InfinyOn
Real-time data integration platform for streaming data between systems. It integrates with MotherDuck for loading data from operational systems, APIs, files, or event streams.
Mage
Mage is a data integration platform for connecting data sources to warehouses. It integrates with MotherDuck for loading data from operational systems, APIs, files, or event streams.
Polytomic
Use Polytomic to sync data to and from MotherDuck for ETL and reverse ETL workflows.
Salesforce
Salesforce is a CRM platform for sales, marketing, service, and customer data. To analyze Salesforce data in MotherDuck, use an ingestion tool that supports Salesforce as a source and MotherDuck as a destination.
Shopify
Load Shopify orders, customers, and products into MotherDuck on a schedule with a Flight that runs dlt's Shopify source.
Sling
Data integration platform for connecting data sources to warehouses. It integrates with MotherDuck for loading data from operational systems, APIs, files, or event streams.
Stacksync
Stacksync helps your teams access and manipulate CRM and ERP data through your existing databases. It integrates with MotherDuck for loading data from operational systems, APIs, files, or event streams.
Streamkap
Streamkap is a stream processing platform built for Change Data Capture (CDC) and event sources. It makes it easy to move operational data into analytics systems like MotherDuck with low latency and high reliability. Streamkap offers various sources, including PostgreSQL, MySQL, SQL Server, a range of SQL and NoSQL databases, Kafka, and other storage systems.
Stripe
Load Stripe customers, subscriptions, invoices, and balance transactions into MotherDuck on a schedule with a Flight that runs dlt's Stripe source.
Unstructured.io
Unstructured.io is an ingestion platform for processing unstructured data. It integrates with MotherDuck for loading data from operational systems, APIs, files, or event streams.