ducta provides a unified framework for constructing data pipelines supporting batch processing, real-time streaming, machine learning workflows, and hybrid scenarios. It includes features for ETL, data quality, orchestration, and integration with Spark and other big-data tools. The package is available on PyPI with an Apache-2.0 license and is primarily used by data engineers and ML engineers to simplify complex data workflow development.
In the Developer Tools space, ducta takes a focused approach. It focuses on building and orchestrating complex data pipelines that span batch, streaming, and machine learning workloads in a single framework. It is built as an open-source project for developers. The project is open source (Apache-2.0). It runs on the command line, and it can be self-hosted.
Faustino Lopez Ramos builds and maintains ducta, and it first shipped in 2026. Development happens publicly on GitHub with 1 commit in the last 90 days. Key capabilities include Data Pipeline, ETL, and streaming.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Same category — not a similarity match