Data Engineer, Pipelines and Cloud Data Warehousing
[email protected] · (404) 555-0129 · Atlanta
Summary
Data engineer with four years building batch and streaming pipelines on AWS for a freight company tracking 2 billion shipment events a month. Works in Python, SQL, Spark, Airflow and dbt, with a focus on pipelines that run on time and data people can trust. AWS Certified Data Engineer. Seeking a senior data engineer role in financial technology.
Experience
- Built a streaming pipeline with Kafka and Spark that loads 2 billion truck and package tracking events a month into Snowflake within 5 minutes of each scan.
- Rebuilt 70 nightly jobs in Airflow with retries and alerts, raising on-time pipeline runs from 91 to 99.5 percent.
- Added automated data quality checks to 40 dbt models, catching bad partner files before they reached the finance reports.
- Cut the monthly Snowflake and AWS bill by $8,000 ($96,000 a year) by clustering large tables and shutting down idle compute.
- Moved 25 reports from hand-run SQL scripts to scheduled jobs, freeing 2 analysts for about 15 hours a week.
Skills
Certifications
AWS Certified Data Engineer, Associate, Amazon Web Services (2025)
Databricks Certified Data Engineer Associate, Databricks (2024)
