<?xml version="1.0" encoding="UTF-8"?>
<feed xmlns="http://www.w3.org/2005/Atom">
    <title>etl</title>
    <link rel="self" type="application/atom+xml" href="https://links.biapy.com/guest/tags/928/feed"/>
    <updated>2026-08-01T20:09:38+00:00</updated>
    <id>https://links.biapy.com/guest/tags/928/feed</id>
            <entry>
            <id>https://links.biapy.com/links/12783</id>
            <title type="text"><![CDATA[Talaxie]]></title>
            <link rel="alternate" href="https://talaxie.deilink.fr/#technical" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/12783"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[The open source future of data integration

Talaxie is a community fork of Talend Open Studio (DI, ESB, BD), aiming to ensure the continuity, stability, and scalability of open source ETL tools despite Talend&amp;#039;s discontinuation.]]>
            </summary>
            <updated>2026-05-18T09:10:01+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/11349</id>
            <title type="text"><![CDATA[SQLFluff]]></title>
            <link rel="alternate" href="https://www.sqlfluff.com/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/11349"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[A modular SQL linter and auto-formatter with support for multiple dialects and templated code. 

SQLFluff is an open source, dialect-flexible and configurable SQL linter. Designed with ELT applications in mind, SQLFluff also works with Jinja templating and dbt. SQLFluff will auto-fix most linting errors, allowing you to focus your time on what matters.

- [SQLFluff @ GitHub](https://github.com/sqlfluff/sqlfluff).]]>
            </summary>
            <updated>2025-12-31T12:45:44+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/11155</id>
            <title type="text"><![CDATA[Singer]]></title>
            <link rel="alternate" href="https://www.singer.io/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/11155"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Simple, Composable, Open Source ETL

Singer powers data extraction and consolidation for all of your organization’s tools.

- [Singer @ GitHub](https://github.com/singer-io).]]>
            </summary>
            <updated>2025-12-02T16:18:00+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/11154</id>
            <title type="text"><![CDATA[Meltano]]></title>
            <link rel="alternate" href="https://meltano.com/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/11154"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Extract &amp;amp; Load with joy.

 Meltano: the declarative code-first data integration engine that powers your wildest data and ML-powered product ideas. Say goodbye to writing, maintaining, and scaling your own API integrations. 

- [Meltano @ GitHub](https://github.com/meltano/meltano).

Related contents:

- [Taming the Data Sources: A Scalable Extraction Strategy with Meltano @ Blueprintdata](https://blueprintdata.xyz/blog/modern-data-stack-meltano).]]>
            </summary>
            <updated>2025-12-02T16:15:26+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/10846</id>
            <title type="text"><![CDATA[Bruin]]></title>
            <link rel="alternate" href="https://getbruin.com/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/10846"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Your last data platform.
Reliable data. 10x faster, 90% less complexity.

 Build data pipelines with SQL and Python, ingest data from different sources, add quality checks, and build end-to-end flows. 

Bruin is a data pipeline tool that brings together data ingestion, data transformation with SQL &amp;amp; Python, and data quality into a single framework. It works with all the major data platforms and runs on your local machine, an EC2 instance, or GitHub Actions.

- [Bruin @ GitHub](https://github.com/bruin-data/bruin).

Related contents:

- [Digest #186: Inside the AWS Outage, Docker Compose in Production, F1 Hacks and 86,000 npm Packages Attacks @ DevOps Bulletin](https://www.devopsbulletin.com/p/digest-186-inside-the-aws-outage).]]>
            </summary>
            <updated>2025-11-03T10:19:09+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/951</id>
            <title type="text"><![CDATA[Mage AI]]></title>
            <link rel="alternate" href="https://www.mage.ai/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/951"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Magical Data Engineering Workflows.

 🧙 Build, run, and manage data pipelines for integrating and transforming data. 

Mage is a hybrid framework for transforming and integrating data. It combines the best of both worlds: the flexibility of notebooks with the rigor of modular code.

- [Mage AI @ GitHub](https://github.com/mage-ai/mage-ai).

Related contents:

- [Alternatives to Talend – How To Migrate Away From Talend For Your Data Pipelines @ Seattle Data Guy](https://www.theseattledataguy.com/alternatives-to-talend-how-to-migrate-away-from-talend-for-your-data-pipelines/).]]>
            </summary>
            <updated>2025-08-28T18:36:35+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/1462</id>
            <title type="text"><![CDATA[PeerDB]]></title>
            <link rel="alternate" href="https://www.peerdb.io/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/1462"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Fast, Simple and a cost effective tool to replicate data from Postgres to Data Warehouses, Queues and Storage.

PeerDB is an ETL/ELT tool built for PostgreSQL. We implement multiple Postgres native and infrastructural optimizations to provide a fast, reliable and a feature-rich experience for moving data in/out of PostgreSQL.

- [PeerDB @ GitHub](https://github.com/PeerDB-io/peerdb).

Related contents:

- [Reliably Replicating Data Between PostgreSQL and ClickHouse Part 1 - PeerDB Open Source @ BenjaminWootton.com](https://benjaminwootton.com/insights/clickhouse-peerdb-cdc/).
- [Postgres Is the Gateway Drug @ Vignesh Ravichandran](https://viggy28.dev/article/postgres-gateway-drug/).]]>
            </summary>
            <updated>2026-03-23T16:36:55+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/1971</id>
            <title type="text"><![CDATA[Pyper]]></title>
            <link rel="alternate" href="https://github.com/pyper-dev/pyper" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/1971"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Concurrent Python made simple.

Pyper is a flexible framework for concurrent and parallel data-processing, based on functional programming patterns. Used for 🔀 ETL Systems, ⚙️ Data Microservices, and 🌐 Data Collection]]>
            </summary>
            <updated>2025-08-28T21:24:22+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/2111</id>
            <title type="text"><![CDATA[Airbyte]]></title>
            <link rel="alternate" href="https://airbyte.com/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/2111"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Open-Source Data Movement for LLMs. AI Platform.
Data integration platform for ELT pipelines from APIs, databases &amp;amp; files to databases, warehouses &amp;amp; lakes.

The leading data integration platform for ETL / ELT data pipelines from APIs, databases &amp;amp; files to data warehouses, data lakes &amp;amp; data lakehouses. Both self-hosted and Cloud-hosted. 

- [Airbyte @ GitHub](https://github.com/airbytehq/airbyte).]]>
            </summary>
            <updated>2026-04-22T12:36:02+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/2208</id>
            <title type="text"><![CDATA[Pathway]]></title>
            <link rel="alternate" href="https://pathway.com/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/2208"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Power Your AI with Live Data.

Python ETL framework for stream processing, real-time analytics, LLM pipelines, and RAG. 

- [Pathway @ GitHub](https://github.com/pathwaycom/pathway).]]>
            </summary>
            <updated>2025-09-10T11:29:30+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/7027</id>
            <title type="text"><![CDATA[Apache Airflow]]></title>
            <link rel="alternate" href="https://airflow.apache.org/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/7027"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Airflow is a platform created by the community to programmatically author, schedule and monitor workflows.

- [Apache Airflow @ GitHub](https://github.com/apache/airflow/).

Related contents:

- [Apache Airflow Configuration and Tuning @ DZone](https://dzone.com/articles/apache-airflow-configuration-and-tuning).
- [Exploring Apache Airflow for Batch Processing Scenario @ DZone](https://dzone.com/articles/exploring-apache-airflow-for-batch-processing-scen).
- [What Is Apache Airflow @ Seattle Data Guy](https://www.theseattledataguy.com/what-is-apache-airflow-data-engineering-consulting/).
- [Alternatives to Talend – How To Migrate Away From Talend For Your Data Pipelines @ Seattle Data Guy](https://www.theseattledataguy.com/alternatives-to-talend-how-to-migrate-away-from-talend-for-your-data-pipelines/).
- [Improving workflow orchestration with Apache Airflow 3.1 in Cloud Composer @ Google Cloud Blog](https://cloud.google.com/blog/products/data-analytics/cloud-composer-supports-apache-airflow-31/).
- [How You can Automate PostgreSQL Backups to S3 Using Apache Airflow! @ Kube Blogs](https://www.kubeblogs.com/automate-postgresql-backups-to-s3/).
- [Générer un Excel volumineux avec Airflow sans dépassement de mémoire @ .LOUD :fr:](https://loud-technology.com/blog/export-excel-streaming-airflow/).]]>
            </summary>
            <updated>2026-07-15T07:06:01+00:00</updated>
        </entry>
    </feed>
