Hey, I'm Gabriel León Castro
Where data guides, decisions thrive
Analytics Engineer with 5+ years building scalable data models, semantic layers and BI solutions.
- Snowflake
- SQL
- Python
- Power BI
- San José, Costa Rica
About
Analytics Engineer with 5+ years building scalable data models, semantic layers, and BI solutions on modern stacks (Snowflake, Power BI, SQL-based ELT). Proven track record designing production-ready reporting solutions, defining business metrics, and performing deep-dive analysis to identify root causes and improvement opportunities. Experienced translating business requirements into analytical solutions and dashboards that support operational decision-making across cross-functional teams.
With a Data Science degree from LEAD University and a SnowPro Core certification, my approach is rooted in precision and a passion for continuous improvement — turning messy data into models that people actually trust and use.
Experience
-
Consultant II, Analytics Engineering @ Hakkōda, an IBM Company
- Redesigned Power BI semantic models that cut dataset refresh times by ~42% (from ~60 to ~35 min), improving usability for 6 business stakeholders.
- Developed and supported Snowflake-based data warehouse solutions, including data modeling, ingestion pipelines, and transformation layers aligned with analytics best practices.
- Collaborated with cross-functional teams (data engineers, analysts, and business users) to gather requirements and deliver scalable, production-ready analytics solutions.
-
Data Analyst @ SGF Global
- Developed and maintained dashboards and reports using Power BI and Tableau.
- Designed SQL queries and scripts to extract, transform, and analyze data from relational databases.
- Built forms and automations using Power Apps and Power Automate to improve operational workflows.
- Built automated reporting and exploratory analysis in Python/Pandas on survey and operational datasets.
-
Data Specialist @ DHL
- Designed and implemented a centralized database infrastructure for the Americas Service Desk.
- Built a Gradient Boosting classifier forecasting Service Desk demand a quarter ahead, deployed on Azure as an alerting system; widening the training history lifted accuracy from 63% to 81%.
- Analyzed agent KPIs to support performance improvement and compensation models.
- Built dashboards and reports using Power BI and Python to communicate insights to stakeholders.
- Cleaned, transformed, and analyzed large datasets using SQL.
-
Data Analyst @ Intel Corporation
- Developed SQL queries and Python scripts to automate data collection and analysis.
- Conducted exploratory data analysis to identify trends and performance improvement opportunities.
- Built dashboards and reports using Power BI and Excel for leadership teams.
- Modeled and analyzed processor architecture data to support current architecture optimization and future architecture development.
Projects
Client work can't be shown, so these are written as case studies: the problem, what I built, and what changed as a result.
-
Forecasting Service Desk demand
DHL · 2021 - 2022
63% → 81%
prediction accuracy
- The problem
- Staffing at the Americas Service Desk was purely reactive. When call and chat volume rose the team hired agents, and when it fell they let them go. Nobody could see a surge coming, so every swing in demand became a hiring or firing decision made under pressure.
- What I did
- A Gradient Boosting classifier predicting, one quarter ahead, whether incoming volume would exceed what the existing team could absorb. The threshold came from the agents’ actual daily capacity, so the output mapped onto a staffing decision rather than an abstract number. The first version trained on 2021 alone and reached 63%; extending the window back to 2018 lifted it to 81%, because the wider history exposed the model to both pre-pandemic and pandemic demand instead of a single regime.
- The outcome
- Deployed on Azure as an alerting system: when the model projected that a team lead’s group would struggle to meet volume, an automated email went out and we reviewed it together. Team leads moved from reacting to volume swings to planning staffing weeks ahead on a data-backed basis.
- Python
- scikit-learn
- Gradient Boosting
- Azure
- SQL
-
Rebuilding reporting on Snowflake
Hakkōda, an IBM Company · 2024 - 2026
~42%
faster dataset refresh
- The problem
- Reporting ran on Microsoft SQL Server while the client moved its warehouse to Snowflake. The Power BI semantic models refreshed several times a day at roughly 60 minutes each, so refresh cycles collided — a new one starting before the previous had finished. The models carried unused columns, bidirectional relationships, heavy Power Query transformations, and DAX doing work that belonged upstream.
- What I did
- Worked alongside the data engineer to migrate the warehouse to Snowflake, using Snowpipe for ingestion and SQL for the transformation layer, and modeled the star schema underneath it. With the warehouse in place I pushed the transformations and calculations down into it, so they run once against Snowflake instead of on every refresh, and rebuilt the semantic models on top of the star schema.
- The outcome
- Refresh time fell from ~60 to ~35 minutes. The refresh cycles stopped overlapping, and the six business stakeholders who rely on these models stopped competing with the pipeline for their own data.
- Snowflake
- Snowpipe
- SQL
- Power BI
- Star schema
-
Modeling processors before the silicon exists
Intel Corporation · 2019 - 2021
- The problem
- Processor architecture decisions are made long before there is any hardware to measure, so they rest entirely on what simulation can be made to reveal.
- What I did
- Analyzed simulation output, built new components into the processor model itself, and added performance counters to capture behaviour that was not being measured. The work spanned Python, C++ and SQL over large simulation datasets, with the reporting layer built in Power BI and Streamlit so results could be explored rather than only read.
- The outcome
- The analysis fed directly to the principal engineer leading the architecture, informing both optimization of the current design and decisions on future ones.
- Python
- C++
- SQL
- Big Data
- Power BI
- Streamlit
Skills
Data & Analytics
Data Engineering
Programming & Libraries
Databases
Tools & Productivity
Also familiar with
Education & Certifications
Education
-
Bachelor's Degree in Data Science
LEAD University
January 2020 - December 2023 · San José, Costa Rica
-
Technical Degree in Computer Networking
Colegio Técnico Profesional de Hatillo
February 2017 - December 2019 · San José, Costa Rica
Certifications
-
SnowPro Core Certified
Snowflake
Issued January 2025 · Valid through January 2027