I'm a Data Engineer with 4+ years building large-scale batch pipelines, ETL/ELT systems, and lakehouse architectures β the infrastructure that data science and AI teams depend on. My work spans the full data lifecycle: ingesting from source systems, transforming with Spark, and serving reliable, governed data for analytics and machine learning.
- π Currently building data pipelines and migrating legacy warehouses onto the Databricks Lakehouse
- π§ Interested in the intersection of data engineering, applied machine learning, and AI-driven data platforms
- π Co-author of a Scopus-indexed publication on ensemble machine learning for imbalanced bioassay data
- π¬ Always happy to talk data pipelines, Spark, or applied ML β feel free to reach out
- π« Reach me at [email protected]
Languages & Data Science
Big Data & Processing
Cloud & Lakehouse
Collaboration & Workflow
Engineering & Tools
Thanks for stopping by β let's connect.


