Skip to content
#

auto-loader

Here are 27 public repositories matching this topic...

End-to-end Azure Databricks retail data engineering project using Medallion Architecture (Bronze, Silver, Gold). Implements Auto Loader, Unity Catalog, Delta Lake, SCD Type 1 & 2 dimensions, and Fact Orders for analytics-ready star schema modeling.

  • Updated Jan 24, 2026

A production-shaped Databricks lakehouse pipeline that ingests, classifies, quality-gates, and aggregates core-banking transaction data. Built on the medallion architecture with Auto Loader, Delta Lake, Unity Catalog, enforced data quality, and infrastructure-as-code deployment via Databricks Asset Bundles.

  • Updated Sep 28, 2026
  • Python

⚡ Real-time sales analytics pipeline using PySpark Structured Streaming on Databricks Free Edition — Auto Loader, windowed aggregations, watermarking, and Delta Lake sink. Beginner-friendly with full README.

  • Updated Jun 5, 2026
  • Jupyter Notebook

Databricks lakehouse on Serverless: medallion pipeline (Auto Loader → Bronze/Silver/Gold), SCD2 dimensions, data-quality quarantine and reconciliation, Slack alerts, Liquid Clustering. 288 unit tests + end-to-end integration test. CI/CD with Asset Bundles, GitHub Actions and service principals. 3M rows in 11m38s.

  • Updated Oct 7, 2026
  • Python

Streaming fraud detection on Databricks: Auto Loader ingest into a Unity Catalog medallion, verified end to end (v0.1.0); Silver, ML scoring and alerts next. PaySim data.

  • Updated Oct 6, 2026
  • Python

Add this topic to your repo

To associate your repository with the auto-loader topic, visit your repo's landing page and select "manage topics."

Learn more