Skip to content
View Mahmoud2saad's full-sized avatar
🎯
Focusing
🎯
Focusing

Block or report Mahmoud2saad

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Mahmoud2saad/README.md

Data Engineer · Data Warehousing & Analytics Engineering



Typing SVG

👋 About Me

I turn messy, multi-source data into pipelines and dashboards that people actually trust and use. I design for the 3am pager alert, not just the demo — with tested SCD2 dimensional models, automated data quality gates, and disaster-recovery runbooks — while keeping the end report or dashboard the whole point of the pipeline, not an afterthought.

Currently seeking Data Engineer / Analytics Engineer roles. 3 production-grade pipelines shipped with CI, 26+ automated tests, and dbt/Airflow/Kafka experience across both cloud and local-hybrid stacks.


💼 Experience

Data Engineering & Analytics Intern National Telecommunication Institute (NTI) · April 2026 – July 2026

  • Designed and shipped 3 interactive Power BI dashboards tracking customer-segment KPIs and telecom performance, adopted for weekly use by operations and management teams
  • Analyzed 100K+ records across multi-source retail and telecom data (SQL, Pandas, NumPy) to surface behavioral patterns that shaped the current quarter's campaign targeting strategy
  • Ran A/B testing frameworks to measure segment lift — results were adopted as the marketing team's campaign baseline
  • Automated recurring KPI reporting (Python + Power Query/PivotTables), cutting manual reporting time by 4+ hours per cycle

🎯 Core Strengths

🔧 Pipeline Engineering Airflow, Kafka, and PySpark across batch and real-time workloads
🏛️ Dimensional Modeling Star schema, SCD Type 1/2, in SQL Server, Snowflake, and DuckDB
Data Quality & Testing dbt tests, automated DQ frameworks that fail builds on critical issues
📊 The Last Mile Power BI dashboards and DAX measures business users actually rely on
🛡️ Production Reliability Documented SLOs, tested DR runbooks, access-control design

🛠️ Technical Skills

Category Tools
Orchestration & Streaming Apache Airflow · Apache Kafka · Debezium (CDC)
Processing & Transformation PySpark · dbt · Python (Pandas, NumPy)
Warehousing & Modeling SQL Server · Snowflake · DuckDB · Star Schema · SCD Type 1 & 2 · Medallion Architecture
BI & Reporting Power BI · DAX · SSAS (OLAP Cubes) · Excel
Infrastructure & Monitoring Docker · Grafana
Languages Python · T-SQL · DAX

🏗️ Featured Projects

Airflow · Kafka · dbt · PySpark · Databricks · Snowflake

  • Hybrid platform: identical code runs locally (DuckDB/PySpark) or in the cloud (Databricks/ADLS/Snowflake) — one environment variable switches modes
  • Real-time Kafka streaming + CDC replication feeding a dbt-built Medallion (Bronze/Silver/Gold) dimensional model
  • Shipped with SLOs, a tested disaster-recovery runbook, and a PAN-tokenization security design — the ops layer most portfolio projects skip

Airflow · PySpark · PostgreSQL · Snowflake

  • Bronze/Silver/Gold pipeline with watermark-based incremental extraction and SCD Type 2 history
  • Automated data quality framework that fails the build on critical checks, not just logs them
  • 26 unit + integration tests running in GitHub Actions CI, with bronze load and DQ checks on a live Airflow schedule

Snowflake · dbt · Airflow · Kafka

  • Real-time CDC ingestion and Kafka streaming pipelines feeding dbt-modeled, BI-ready marts

Power BI · DAX · Power Query

  • Full BI workflow — data cleaning, star-schema modeling, 11+ custom DAX measures — into an interactive 3-page dashboard covering attrition, attendance, and performance
More projects — Sales Analytics DW · Hotel Booking DW (SSIS/SSAS, SCD2, OLAP cubes)

SSIS · SSAS · SQL Server · Power BI — Star-schema warehouse with SCD Type 2 tracking and a multidimensional OLAP cube powering enterprise-wide sales reporting.

SSIS · SSAS · Power BI — End-to-end BI pipeline with SCD Type 2 historical tracking, OLAP cube modeling, and automated ETL for booking and occupancy analytics.


📊 GitHub Stats


📜 Certifications

  • Data Engineer in Python — DataCamp
  • Google Data Analytics Professional Certificate
  • IBM Data Science Professional Certificate
  • Data Analyst in Power BI — DataCamp
  • Data Analyst in Python — DataCamp

📫 Open to Data Engineer / Analytics Engineer roles

If you're hiring or just want to talk pipelines, my inbox is open.


Pinned Loading

  1. End_To_End_Banking_Pipeline End_To_End_Banking_Pipeline Public

    Hybrid batch/streaming banking pipeline with a tested DR runbook, documented SLOs, and PAN-tokenization-first security design. Medallion architecture, Kafka/CDC, dbt star schema, Airflow orchestrat…

    Python

  2. Hotel-Booking-DataWarehouse Hotel-Booking-DataWarehouse Public

    Full BI pipeline with SSIS ETL, SSAS OLAP cube, SCD Type 2, and Power BI drill-through dashboards

    HTML 1

  3. Sales-Analytics-DataWarehouse Sales-Analytics-DataWarehouse Public

    Star schema DWH with SSIS ETL, SSAS OLAP cube, SCD Type 1 & 2, and Power BI dashboards

    TSQL

  4. Airline-Loyalty-Data-Warehouse Airline-Loyalty-Data-Warehouse Public

    3-layer ETL pipeline (ODS→STG→DWH) using SSIS with SSAS cube and Power BI loyalty analytics

    TSQL

  5. Books-Data-Warehouse-End-to-End Books-Data-Warehouse-End-to-End Public

    End-to-End Data Warehouse & Business Intelligence Solution for a Bookstore using SQL Server, SSIS, SSAS, Snowflake Schema, and Power BI.

  6. Therapist-Location-recommender- Therapist-Location-recommender- Public

    Discovering Excellence: Your Guide to Top Doctors

    Jupyter Notebook 1