Skip to content
View AaronVillegas5's full-sized avatar

Highlights

  • Pro

Block or report AaronVillegas5

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
AaronVillegas5/README.md

Hi there, I'm Aaron Kyle Villegas 👋

I am a data professional and recent UC Irvine graduate (B.S. Applied and Computational Mathematics, Data Science specialization) passionate about building robust, end-to-end data pipelines and uncovering actionable business insights.

I am currently relocating to the Northwest Arkansas area and actively seeking roles as a Data Analyst, BI Analyst, or Junior Data Engineer.

🏆 Honors & Recognition

  • Howard Tucker Award in Mathematics (UC Irvine)
  • Cum Laude & Dean's Honor List (Cumulative GPA: 3.92)
  • Datathon Winner: Best Visualization

🛠️ Tech Stack & Infrastructure

  • Languages: Python, SQL, R, C++, Java
  • Data Warehousing & DBs: Snowflake, Google BigQuery, PostgreSQL
  • BI & Visualization: Tableau, PowerBI
  • DevOps & Infrastructure: Docker, Linux, Git, Tailscale, Syncthing

🚀 Featured Projects

Macro Data Pipeline

Engineered an automated Python ETL pipeline designed to handle large-scale time-series data. Sourced and structured over 9 million rows of data from the FRED and Open-Meteo APIs, developing automated data cleaning and deduplication workflows to route high-reliability datasets into PostgreSQL, Snowflake, and BigQuery.

Retail & Sales Analytics

Developed automated reporting models and a Tableau dashboard tracking 9+ core business KPIs using Python and SQL. Uncovered a 69% holiday sales lift and analyzed store-level demand volatility to recommend optimized inventory and supply chain strategies.

Self-Hosted Homelab

Built and currently manage a personal server environment utilizing a Raspberry Pi 5. The architecture features NVMe storage, containerized application deployment via Docker, and secure remote networking utilizing Tailscale and AdGuard Home.

🧗‍♂️ Outside the Terminal

Before diving full-time into data, I spent time studying quantitative finance and trading (Jane Street Summer Program) and worked for years as a Rock Wall Supervisor and Belayer. When I'm not writing SQL or tinkering with my Docker containers, you can usually find me researching personal finance optimization or at the local climbing gym.

📫 Let's Connect

Pinned Loading

  1. environmental-health-risk-map environmental-health-risk-map Public

    Datathon Winner: Best Visualization — Interactive geospatial dashboard using ML to map environmental health risk across Southern California ZIP codes.

    HTML

  2. walmart_data_analysis walmart_data_analysis Public

    End-to-end data analysis of Walmart sales to identify key revenue drivers, seasonal trends, and actionable business insights.

    Jupyter Notebook

  3. macro-data-pipeline macro-data-pipeline Public

    Automated ETL pipeline ingesting macroeconomic and weather APIs into PostgreSQL and Snowflake, utilizing AWS S3 and Python for time-series analytics.

    Python

  4. ABTest-MarketingCampaign ABTest-MarketingCampaign Public

    Statistical A/B test analysis of a 588K-user marketing campaign comparing advertisement vs. PSA conversion rates using hypothesis testing and confidence intervals

    Jupyter Notebook

  5. uci-bank-marketing-analysis uci-bank-marketing-analysis Public

    Data-driven analysis and ML modeling to optimize marketing targeting and improve campaign efficiency.

    Jupyter Notebook