I’m a senior Data Engineer with 9+ years of experience building enterprise-scale data platforms, scalable ETL pipelines, and cloud-native lakehouse architectures on AWS. I design end-to-end pipelines that consolidate structured, semi-structured, and unstructured data into secure, governed data lakes and analytics warehouses—using tools like AWS Glue, EMR Serverless, Lambda, Redshift Serverless, and orchestration with Step Functions.
I’m especially hands-on with PySpark and Spark SQL for high-volume distributed transformations, plus modern lakehouse formats such as Apache Iceberg (and Delta Lake), Parquet, and dbt for reliable, testable analytics layers. I also bring strong governance and compliance experience through AWS Lake Formation and data quality/lineage practices, and I’ve built both batch and real-time streaming pipelines using MSK/Kafka and Kinesis for low-latency processing in regulated domains like banking and healthcare.
Language
Work Experience
Education
Qualifications
Industry Experience
Hire a Data Engineer
We have the best data engineer experts on Twine. Hire a data engineer in Cary today.