Skip to content
View shahidmalik4's full-sized avatar

Block or report shahidmalik4

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
shahidmalik4/README.md

👨‍💻 About Me

I'm a data professional with 4+ years of production experience building the data systems behind analytics and reporting, from data pipelines and modeling to database migrations, automation, and data quality.

As the sole data professional at my company, I've owned data workflows end to end, working across production databases, business systems, and operational processes to build reliable and usable data solutions.

🚀 Production Experience

  • Led the migration of a production database from SQL Server to PostgreSQL, including schema conversion, data validation, reconciliation, and production cutover
  • Designed and maintained layered SQL data pipelines across ingestion → staging → cleansing → modeling for sales and CRM data
  • Automated reporting and data workflows with Python, reducing manual effort by ~35%
  • Implemented data quality and validation checks that contributed to a 4 percentage-point improvement in profit margin and ~20% revenue growth
  • Built data models and reporting datasets supporting revenue, profitability, sales, and supply chain analytics

🔧 Analytics Engineering & Data Engineering

Alongside my production work, I build hands-on projects to deepen my experience with modern analytics and data engineering practices.

dbt · Snowflake · Airflow · Docker · FastAPI · CI/CD

I'm currently focused on Analytics Engineering and Data Engineering, with an interest in building reliable pipelines, well-modeled data, and maintainable data systems.

🛠 Tech Stack

Category Tools
Data SQL · PostgreSQL · SQL Server · Python
Analytics Engineering dbt · Data Modeling · ELT
Orchestration & Cloud Airflow · Snowflake
Engineering Docker · FastAPI · Git · GitHub Actions
Analytics Power BI

Pinned Loading

  1. dbt-airflow-data-pipeline dbt-airflow-data-pipeline Public

    A full analytics workflow simulating a real-world business environment! The project starts with raw transactional data (TPCH dataset) and transforms it into clean, aggregated KPIs, ready for analys…

    Python 5

  2. pyspark-snowflake-dbt-pipeline pyspark-snowflake-dbt-pipeline Public

    This project is a data engineering pipeline leveraging PySpark, Snowflake, Airflow, dbt, and Streamlit to extract, transform, and load millions of records daily. It streamlines data processing, ena…

    Python 2

  3. analytics-pipeline-fastapi-dbt analytics-pipeline-fastapi-dbt Public

    A full-stack data analytics pipeline using DBT, FastAPI, Streamlit and Postgres. Transforms raw data into modeled tables and exposes KPIs via API endpoints, with an interactive dashboard for visual…

    Python 1

  4. aws-glue-stepfunctions-etl aws-glue-stepfunctions-etl Public

    This project automates an ETL pipeline using AWS Glue, S3, Athena, and Step Functions to transform raw Airbnb data. It cleanses, enriches, and organizes the data into separate raw and transformed d…

    Python 1

  5. dbt-snowflake-data-pipeline dbt-snowflake-data-pipeline Public

    The dbt Snowflake Data Pipeline project uses dbt to transform data in Snowflake, creating efficient, scalable data models for analysis. It leverages incremental models to handle large datasets and…

    1

  6. data-platform-forge data-platform-forge Public

    A production-style local data platform built with modern data engineering tools. This project simulates a real-world ELT pipeline — from raw data ingestion through transformation and orchestration …

    Python 1