Skip to content
View rickard-garnau's full-sized avatar
🎯
Focusing
🎯
Focusing

Block or report rickard-garnau

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
rickard-garnau/README.md

Hi there! I'm Rickard

Data Engineering student at STI (2025–2027), looking for a LIA internship 11 January – 28 May 2027 in the Stockholm area.

Before studying I spent 15 years at Citymail Sweden AB, where a wrong address or register has real consequences. That is a big part of why I want data to be right from the start.

Stack

  • Languages: Python (Pandas, OOP) · SQL (CTEs) · dbt · dlt
  • Cloud & Infrastructure: Azure · Terraform · Docker
  • Data Platforms: Databricks · Delta Live Tables · PySpark · Snowflake · Apache Kafka
  • Databases: PostgreSQL · DuckDB
  • Backend & APIs: FastAPI · REST
  • Modeling: ER-modeling · 3NF Normalization · Dimensional modeling · Medallion architecture
  • BI: Power BI · Streamlit · Evidence.dev
  • Practices: Git · GitHub Actions · pytest · Agile/Scrum

Projects

Fullstack app for analysis and visualization of solar and lunar eclipses, based on NASA's Five Millennium Catalogs.

  • FastAPI backend and Streamlit frontend as separate services
  • Containerized with Docker, deployed to Azure with Terraform (Container App and Web App)
  • Error handling for failed API calls, backend URL controlled via environment variable

Medallion pipeline on Databricks for ultra marathon results (7.4M rows, 1990–2022).

  • Streaming ingestion via Delta Live Tables into bronze
  • Silver: unit standardization, date parsing, performance normalization
  • Dimensional model in gold: fct_results, dim_athlete, dim_event and analytical views
  • Genie space for ad hoc questions, verified manually against SQL
  • Databricks dashboard on the gold views

FoodHub (group project)

Recipe search platform: FastAPI, Kafka and PostgreSQL in Docker. Kafka producer/consumer streams data from the Spoonacular API into PostgreSQL (staging → curated), with a cache-first strategy to limit external API calls. ETL with Pydantic validation, NaN handling and fuzzy ingredient matching. Scrum Master for half the project of a team of 5 people.

STHLMs Puls (group project)

Event guide for Stockholm in Power BI, with a map of venues, charts per genre and weekday, and filters on date and subcategory. Data from Ticketmaster, Visit Stockholm, Berns, Fasching and Google Places, plus weather via API. I built the start page, the performing-arts and nightlife pages, and parts of the data pipeline. Also a Streamlit version.

Contact

More projects and course work under my repositories.

Pinned Loading

  1. azure_python_fullstack_lab azure_python_fullstack_lab Public

    Jupyter Notebook

  2. ultra-marathon-pipeline ultra-marathon-pipeline Public

    Ultra marathon analytics pipeline built with Databricks, PySpark and DLT — bronze to gold medallion architecture.

    Jupyter Notebook

  3. data-platform-project data-platform-project Public

    Recipe search platform with fuzzy matching, Kafka streaming and PostgreSQL. Built with FastAPI, Docker and Supabase.

  4. stock-data-pipeline stock-data-pipeline Public

    ELT data pipeline for stock data using FastAPI, PostgreSQL & Pandas

    Python 1

  5. visualization-project-streamlit visualization-project-streamlit Public

    Forked from LisaYllander92/Visualization_project

    A data-driven cultural guide for Stockholm built with Streamlit. Aggregates event data from Ticketmaster, VisitStockholm, Fasching and Berns, combined with real-time weather forecasts from Open-Met…

    Jupyter Notebook