Skip to content
View shahidmalik4's full-sized avatar

Block or report shahidmalik4

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
shahidmalik4/README.md

👨‍💻 About Me

I'm a Senior Data Analyst with nearly 5 years of production experience building and maintaining the data workflows behind analytics and reporting.

As the primary data professional at my company, I've taken ownership of data end to end, from SQL-based data pipelines, data modeling, and reporting to automation, data quality, and production database operations.

Over time, my work has moved beyond traditional analytics toward the engineering side of data. I'm now focused on transitioning into Analytics Engineering and Data Engineering, with a primary interest in building reliable data models, transformation pipelines, and maintainable data systems.

🚀 Production Experience

  • Led the migration of a production database from SQL Server to PostgreSQL, covering schema conversion, data validation, reconciliation, and production cutover
  • Designed and maintained layered SQL data pipelines across ingestion → staging → cleansing → modeling for sales, CRM, and operational data
  • Automated recurring reporting and data workflows using Python and PostgreSQL, including data extraction, validation, PDF report generation, and automated email delivery, significantly reducing manual reporting effort.
  • Implemented automated data quality and validation checks to improve reliability and reduce discrepancies across business reporting
  • Built reusable SQL data models and reporting datasets supporting revenue, profitability, sales, customer, and operational analytics
  • Delivered data-driven improvements that contributed to ~20% revenue growth and a 4 percentage-point improvement in profit margin

🔧 Analytics Engineering & Data Engineering

Alongside my production experience, I build hands-on projects to develop and apply modern Analytics Engineering and Data Engineering practices.

My primary focus is Analytics Engineering — data transformation, modeling, testing, data quality, and building reliable datasets for analytics.

I'm also developing deeper Data Engineering skills around pipeline orchestration, APIs, infrastructure, and production-oriented data workflows.

dbt · Snowflake · Airflow · Docker · FastAPI · CI/CD

🛠 Tech Stack

Category Tools
Data & Databases SQL · PostgreSQL · SQL Server · Python
Analytics Engineering dbt · Data Modeling · ELT · Data Quality
Data Engineering Airflow · Snowflake · Data Pipelines · APIs
Engineering & DevOps Docker · FastAPI · Git · GitHub Actions
Analytics & BI Power BI

Pinned Loading

  1. dbt-airflow-data-pipeline dbt-airflow-data-pipeline Public

    A full analytics workflow simulating a real-world business environment! The project starts with raw transactional data (TPCH dataset) and transforms it into clean, aggregated KPIs, ready for analys…

    Python 5

  2. pyspark-snowflake-dbt-pipeline pyspark-snowflake-dbt-pipeline Public

    This project is a data engineering pipeline leveraging PySpark, Snowflake, Airflow, dbt, and Streamlit to extract, transform, and load millions of records daily. It streamlines data processing, ena…

    Python 2

  3. analytics-pipeline-fastapi-dbt analytics-pipeline-fastapi-dbt Public

    A full-stack data analytics pipeline using DBT, FastAPI, Streamlit and Postgres. Transforms raw data into modeled tables and exposes KPIs via API endpoints, with an interactive dashboard for visual…

    Python 1

  4. aws-glue-stepfunctions-etl aws-glue-stepfunctions-etl Public

    This project automates an ETL pipeline using AWS Glue, S3, Athena, and Step Functions to transform raw Airbnb data. It cleanses, enriches, and organizes the data into separate raw and transformed d…

    Python 1

  5. dbt-snowflake-data-pipeline dbt-snowflake-data-pipeline Public

    The dbt Snowflake Data Pipeline project uses dbt to transform data in Snowflake, creating efficient, scalable data models for analysis. It leverages incremental models to handle large datasets and…

    1

  6. data-platform-forge data-platform-forge Public

    A production-style local data platform built with modern data engineering tools. This project simulates a real-world ELT pipeline — from raw data ingestion through transformation and orchestration …

    Python 1