apache-spark
DLT-META
https://databrickslabs.github.io/dlt-meta/
Metadata driven Spark Declarative Pipelines framework for bronze/silver pipelines.
DLT-META is a metadata-driven framework designed to work with Lakeflow Declarative Pipelines. This framework enables the automation of bronze and silver data pipelines by leveraging metadata recorded in an onboarding JSON file. This file, known as the Dataflowspec, serves as the data flow specification, detailing the source and target metadata required for the pipelines.
Related contents:
Added 7 months ago
Apache Spark
https://spark.apache.org/
Unified Engine for large-scale data analytics.
Apache Spark⢠is a multi-language engine for executing data engineering, data science, and machine learning on single-node machines or clusters.
Related contents:
Added 7 months ago