Data Engineer with 6+ years of experience building scalable batch and streaming data pipelines using Spark, Kafka, and PySpark, handling multi-terabyte datasets across enterprise systems. Experienced in AWS-based data platforms (S3, Redshift, Kinesis), real-time stream processing, and ETL/ELT workflows using SQL, Informatica, and dbt.
Strong background in orchestration with Apache Airflow, data warehousing with Snowflake/Redshift, and data modeling (star schema, dimensional modeling). Focused on data quality and governance using Great Expectations, along with monitoring and reliability improvements through CloudWatch and Grafana, and deploying containerized pipelines with Docker and Kubernetes.
Skills
Language
Work Experience
Education
Qualifications
Industry Experience
Skills
Hire a Data Engineer
We have the best data engineer experts on Twine. Hire a data engineer today.