Data Pipeline
9Companies categorized as Data Pipeline.
Jitsu is an open-source customer data platform that captures event data from websites, apps, and servers and streams it into data warehouses and analytics tools. It is used by developers and data teams to collect and route event data for analysis.
Fivetran provides a platform that moves, manages, and transforms data from various sources into data warehouses, lakes, or applications for analytics, operations, and AI.
Bytewax is an open source Python framework that lets developers build scalable dataflow pipelines for processing any data stream.
WarpStream offers a diskless, Kafka-compatible streaming platform that runs as stateless agents in a VPC and writes data directly to cloud object storage. It lets developers and data teams stream data without managing local disks or inter-AZ networking.
Sequin streams change data from PostgreSQL databases to destinations such as Kafka, Redis, SQS, and other sinks, providing exactly‑once delivery and strict ordering. It is used by developers and data teams to build change‑data‑capture pipelines without managing additional infrastructure.
Streamkap offers a managed change data capture and streaming platform that delivers database changes to AI agents, applications, and data warehouses with sub‑50 ms latency. It is used by developers and data teams to build real‑time data pipelines without managing Kafka or Flink.
Multiwoven is an open-source reverse ETL platform that enables users to sync and activate customer data from data warehouses to business tools on self-hosted infrastructure. It is used by data engineers and analysts to move data from sources such as Redshift, Snowflake, or BigQuery into downstream applications.
PeerDB provides a service that replicates data from PostgreSQL databases to data warehouses, queues, and storage systems. It is used by companies that rely on Postgres as a primary data source.
dltHub is a managed platform that deploys, monitors, and scales data pipelines built with the open-source Python library dlt. It provides runtime, observability, data quality checks, and collaboration tools for data engineering teams.