Skip to content
View felipeandrade91's full-sized avatar

Block or report felipeandrade91

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
felipeandrade91/README.md

Hi, I'm Felipe Andrade

Data Analyst | SQL | PostgreSQL | Python | Power BI | Machine Learning | Data Engineering | Customer Analytics | Analytics Engineering | Causal Inference | Experimentation

I'm a Data Analyst and PhD with over 15 years of experience transforming complex real-world datasets into actionable insights through SQL, Python, Power BI, statistics, and analytical modeling.

My scientific background has strengthened my analytical thinking, hypothesis-driven problem solving, and ability to design reproducible analytical workflows.

Today, I apply these skills to solve business problems in Business Intelligence, Analytics Engineering, Customer Analytics, Machine Learning, Data Engineering, and Causal Inference & Experimentation.

Applying scientific rigor and analytical thinking to solve business problems through modern data analytics.


Tech Stack

Analytics & Business Intelligence

  • SQL
  • PostgreSQL
  • Power BI
  • DAX
  • Power Query (M)
  • Analytics Engineering
  • Customer Analytics
  • Customer Segmentation
  • Customer Lifetime Value (CLV)
  • Star Schema
  • Data Modeling

Programming & Analytics

  • Python
  • Pandas
  • Matplotlib
  • R
  • Statistics
  • A/B Testing
  • Causal Inference
  • Treatment Effect Estimation
  • Double Machine Learning
  • Causal Forest
  • Propensity Score Matching
  • Incrementality Analysis
  • Machine Learning
  • Scikit-learn
  • XGBoost
  • FastAPI
  • REST APIs
  • Docker
  • Docker Compose
  • Model Deployment
  • Time Series Forecasting
  • Predictive Modeling
  • Feature Engineering
  • PySpark
  • Databricks
  • Delta Lake
  • Data Pipelines
  • Medallion Architecture
  • Incremental Processing

Tools

  • Git
  • GitHub
  • DBeaver
  • Excel

Featured Projects

⭐ Customer Analytics for Brazilian E-commerce

End-to-end Analytics Engineering and Business Intelligence project built with PostgreSQL, SQL, and Power BI, featuring a Star Schema data model, SQL semantic layer, business KPIs, and interactive executive dashboards.

🔗 Repository

https://github.com/felipeandrade91/Customer-Analytics-for-Brazilian-E-commerce


⭐ NYC Taxi Data Engineering Platform

End-to-end Data Engineering pipeline built with Databricks, PySpark, and Delta Lake, processing more than 38 million NYC Yellow Taxi trips through a Medallion Architecture. The project demonstrates scalable data ingestion, incremental processing, data quality validation, dimensional modeling, and analytical data preparation.

🔗 Repository

https://github.com/felipeandrade91/nyc-taxi-data-engineering


⭐ Customer Segmentation & Customer Lifetime Value Analytics

Customer Analytics project built with PostgreSQL, SQL, and Python, extending the previous analytical foundation through customer feature engineering, RFM segmentation, Historical Customer Lifetime Value (CLV) analysis, and business-oriented data visualization.

🔗 Repository

https://github.com/felipeandrade91/customer-segmentation-clv


⭐ Customer Churn Prediction

An end-to-end Machine Learning project to predict customer churn using the IBM Telco Customer Churn dataset. The project demonstrates the complete data science workflow, including SQL data preparation, exploratory data analysis, feature engineering, predictive modeling, model evaluation and business interpretation.

https://github.com/felipeandrade91/customer-churn-prediction


⭐ Customer Churn Prediction API

Containerized REST API for customer churn prediction using FastAPI, Scikit-learn, Docker, and Docker Compose. The project demonstrates model deployment, input validation, automated testing, and containerized inference using the trained machine learning pipeline from the Customer Churn Prediction project.

🔗 Repository

https://github.com/felipeandrade91/customer-churn-api


⭐ Causal Inference & Experimentation

A practical causal inference project evaluating the incremental impact of digital advertising using A/B Testing, Propensity Score Matching, Inverse Probability Weighting, Double Machine Learning, and Causal Forests. The project combines experimental and observational approaches to estimate treatment effects and investigate how advertising effectiveness varies across users.

🔗 Repository

https://github.com/felipeandrade91/causal-inference-experimentation


⭐ Sales Forecasting with Machine Learning - Rossmann Stores

An end-to-end machine learning project for retail sales forecasting using the Rossmann Store Sales dataset. The project covers exploratory time series analysis, temporal feature engineering, regression modeling, and model interpretation using Linear Regression, Random Forest, and XGBoost.

🔗 Repository

https://github.com/felipeandrade91/sales-forecasting-rossmann


📂 Data Analytics Portfolio

A curated collection of my Data Analytics, Business Intelligence, Analytics Engineering, Data Engineering, Machine Learning, Customer Analytics, SQL, Python, and Power BI projects.

🔗 Repository

https://github.com/felipeandrade91/Data-Analytics-Portfolio


Scientific Background

  • PhD in Animal Biology (UNICAMP)
  • Postdoctoral Researcher (USP)
  • 28 peer-reviewed scientific publications
  • Description of 13 new amphibian species
  • 15+ years working with complex real-world datasets

Connect with Me

💼 LinkedIn

https://linkedin.com/in/felipeandrade91

Pinned Loading

  1. Brazilian-Anuran-Diversity-Dashboard Brazilian-Anuran-Diversity-Dashboard Public

    Interactive Power BI dashboard exploring the spatial and temporal patterns of Brazilian anuran diversity using GBIF occurrence records.

    1

  2. gbif-amphibians-etl-pipeline gbif-amphibians-etl-pipeline Public

    This project was developed as a portfolio piece focused on biodiversity data engineering, demonstrating skills in SQL-based ETL pipelines, data quality assessment, and analytical dataset construction.

    1

  3. GBIF-Brazilian-Amphibian-Biodiversity-Analysis GBIF-Brazilian-Amphibian-Biodiversity-Analysis Public

    This project presents an exploratory and statistical analysis of amphibian occurrence records in Brazil using data from the Global Biodiversity Information Facility (GBIF).

    Jupyter Notebook 1

  4. Customer-Analytics-for-Brazilian-E-commerce Customer-Analytics-for-Brazilian-E-commerce Public

    End-to-end Customer Analytics project using PostgreSQL and Power BI, featuring Analytics Engineering, Star Schema modeling, SQL semantic layers, and interactive business dashboards built from the B…

    1