Jul 2024 – Apr 2026
Senior Data Engineer
Most recentJefferson Bank · San Antonio, TX
- Architected and optimized large-scale Spark SQL pipelines in Databricks, migrating legacy MapReduce workloads to PySpark for significant throughput gains.
- Built ADF pipelines ingesting data from relational and unstructured sources into Azure Data Lake, Databricks, and Azure SQL DW with full lineage tracking.
- Designed Star and Snowflake schemas for the enterprise data warehouse; implemented SCD patterns to preserve historical accuracy across dimension tables.
- Set up end-to-end observability using Prometheus and Grafana across all data pipelines, reducing mean time to detect pipeline failures.
- Deployed Apache Airflow for authoring, scheduling, and monitoring production DAGs; used Kafka for high-throughput web log aggregation feeding downstream analytics.
- Delivered Power BI and Tableau dashboards enabling business teams to compare legacy vs. current data and surface KPIs in real time.