I'm a data engineer with 5+ years of experience designing and maintaining scalable ETL/ELT pipelines, big data architectures, and cloud-based data solutions for enterprise systems. I specialize in Hadoop, Spark, Databricks, and Snowflake for distributed data processing, analytics, and performance optimization. I'm proficient in Python, SQL, and Scala, and have hands-on experience with AWS, Azure, and GCP. I thrive on collaborating with data science and business teams to ensure data reliability, governance, and timely insights. I've led cloud migration projects (Teradata to Snowflake, Oracle to AWS S3) and delivered automated CI/CD for data pipelines using Jenkins and Docker, with a strong focus on quality and security.…

Keerthi Narina

I'm a data engineer with 5+ years of experience designing and maintaining scalable ETL/ELT pipelines, big data architectures, and cloud-based data solutions for enterprise systems. I specialize in Hadoop, Spark, Databricks, and Snowflake for distributed data processing, analytics, and performance optimization. I'm proficient in Python, SQL, and Scala, and have hands-on experience with AWS, Azure, and GCP. I thrive on collaborating with data science and business teams to ensure data reliability, governance, and timely insights. I've led cloud migration projects (Teradata to Snowflake, Oracle to AWS S3) and delivered automated CI/CD for data pipelines using Jenkins and Docker, with a strong focus on quality and security.…

Available to hire

I’m a data engineer with 5+ years of experience designing and maintaining scalable ETL/ELT pipelines, big data architectures, and cloud-based data solutions for enterprise systems. I specialize in Hadoop, Spark, Databricks, and Snowflake for distributed data processing, analytics, and performance optimization. I’m proficient in Python, SQL, and Scala, and have hands-on experience with AWS, Azure, and GCP.

I thrive on collaborating with data science and business teams to ensure data reliability, governance, and timely insights. I’ve led cloud migration projects (Teradata to Snowflake, Oracle to AWS S3) and delivered automated CI/CD for data pipelines using Jenkins and Docker, with a strong focus on quality and security.

See more

Experience Level

Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Intermediate
Intermediate
See more

Work Experience

Data Engineer at BMO Harris Bank
July 1, 2024 - November 6, 2025
Designed and implemented a scalable microservices-based architecture for a financial reporting platform using Apache Spark (Scala) and Snowflake. Built end-to-end ETL/ELT pipelines in AWS Glue to ingest customer transaction data from S3, transform with PySpark, and load into Snowflake for regulatory analytics. Built and scheduled Airflow workflows for daily data loads, enabled rapid failure detection, and migrated on-prem data from Teradata to Snowflake with automated schema mapping and validation. Implemented API-based ingestion, processed unstructured payloads (JSON, CSV, Avro), containerized services with Docker, and deployed on Kubernetes. Established logging with Log4j and CloudWatch, and automated CI/CD with Jenkins and GitLab. Created Power BI data cubes via Azure Analysis Services to support finance reporting.
Data Engineer at Moderna
July 1, 2023 - July 1, 2023
Built cloud-native ETL/ELT pipelines on AWS and Azure Data Factory to ingest and transform high-volume genomic and clinical datasets. Developed Spark jobs on AWS EMR and Databricks, and modeled data in Redshift and S3, enabling scalable analytics for R&D and regulatory reporting. Migrated Informatica flows to Databricks notebooks, implemented end-to-end data ingestion for CSV/JSON/Parquet/ORC, and maintained logging/monitoring with CloudWatch and Datadog. Containerized Spark tasks with Docker and Kubernetes, optimized Spark configs and query plans, and supported AI/ML data preparation workstreams.
Data Engineer at AT&T
July 1, 2021 - July 1, 2021
Designed enterprise data architectures using Azure Data Lake, Synapse, Delta Lake, and Databricks to support scalable telecom data ingestion and processing. Developed PySpark/Python ETL/ELT pipelines with Azure Data Factory, processing petabytes of data with optimized performance and cost. Built SQL transformations and data models in Synapse for BI reporting, and integrated real-time streaming with Kafka along with batch automation. Implemented REST API data delivery, managed Snowflake and Palantir Foundry for analytics, and established data quality checks with Great Expectations and SQL assertions. Set up CI/CD pipelines with Jenkins and Docker, enforced IAM and encryption with data masking, and supported AI/ML feature engineering for model deployment.

Education

Master’s in computer science at Eastern Illinois University, Illinois
January 11, 2030 - November 6, 2025

Qualifications

Add your qualifications or awards here.

Industry Experience

Financial Services, Software & Internet, Telecommunications, Healthcare, Life Sciences, Professional Services