A production-grade data engineering platform implementing the Bronze → Silver → Gold medallion architecture for scalable data ingestion, transformation, quality, governance, and analytics.
-
Updated
Aug 23, 2026 - Python
A production-grade data engineering platform implementing the Bronze → Silver → Gold medallion architecture for scalable data ingestion, transformation, quality, governance, and analytics.
Metadata-driven framework for Databricks Spark Declarative Pipelines. Config-driven, pattern based approach to batch & streaming across the medallion architecture. Deploys via Declarative Automation Bundles. Built for simplicity, extensibility, and alignment with the Databricks product roadmap.
Declarative Implementation of Medallion Architecture - using Opinionated Open-source libraries at different Layers
71M Steam reviews × IGDB in a Microsoft Fabric medallion lakehouse, measuring where review sentiment disagrees with the thumbs.
Azure-oriented geospatial data platform standardizing Canadian climate, hydrometric, wildfire, and disaster datasets into validated spatial risk products for Alberta and British Columbia, with Snowflake, dbt, Power BI, and MapLibre GL.
Academic data platform with an orchestrated warehouse, a natural-language assistant that answers questions with safe SQL, and observability across the whole pipeline.
⬡ TRUSTINT — Trust Intelligence Architecture · provenance-first substrate for trust governance · SQLite (WAL+FTS5) · explicit db_path · idempotent ingest · ADR-driven · collapse-aware · 🥉 v0.1.1
Full-featured Apache Iceberg lakehouse data engineering lab with 19 Spark scenarios in Scala and PySpark, 2 CI-verified Maven Scala Spark applications built by Jenkins and orchestrated by Airflow, Trino SQL and BI, Redpanda streaming, and a medallion architecture across bronze, silver, and gold layers using Docker Compose to run the Atlas platform.
To associate your repository with the medallion topic, visit your repo's landing page and select "manage topics."