Available to hire
Senior Data Engineer with 8+ years of experience architecting enterprise-grade data platforms across banking, financial services, retail, and e-commerce. Expertise includes end-to-end ETL/ELT pipelines, real-time streaming architectures, cloud-native lakehouse solutions, and data governance to support compliance and reliability.
Delivered systems processing billions of events daily, reducing infrastructure costs by $1.2M+ and enforcing zero-incident GDPR/OSFI compliance across audited pipelines. Strong background in Python, Scala, Apache Spark, Kafka, Airflow, Snowflake/Redshift/BigQuery, plus DataOps, IaC, and mentoring data engineering teams.
Experience Level
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Intermediate
Intermediate
Work Experience
Senior Data Engineer at Capital One Canada
April 1, 2025 - PresentDeveloped an AWS-based lakehouse architecture using S3, EMR, Glue, Redshift, and Delta Lake to centralize enterprise data and improve reporting performance by 40%. Built ETL/ELT workflows using PySpark and Airflow for structured, semi-structured, and streaming ingestion from APIs, databases, and Kafka. Implemented real-time Kafka + Spark Streaming pipelines for fraud detection, reducing processing latency from minutes to seconds. Optimized Spark jobs (partitioning, broadcast joins, caching, executor tuning) to reduce processing time by 35%. Designed analytics data models (star schema, Snowflake) and implemented data quality, lineage tracking, RBAC controls, PII masking, and encryption standards. Built Terraform modules and CI/CD pipelines (Jenkins/GitHub Actions), and containerized applications with Docker and Kubernetes for portable deployments.
Data Migration Engineer at KPMG Canada
January 1, 2024 - March 31, 2025Executed enterprise-scale migration of Hadoop, Oracle, SQL Server, and legacy warehouse platforms to AWS cloud using S3, Snowflake, EMR, Glue, and Redshift with minimal downtime. Built and optimized scalable ETL/ELT pipelines using Python, Spark, dbt, and Airflow for automated ingestion, transformation, cleansing, reconciliation, and source-to-target mapping. Implemented data quality and cloud automation frameworks using Great Expectations and Terraform/CloudFormation, improving validation accuracy and processing performance.
Streaming Data Engineer at Canadian Tire Corporation
May 1, 2022 - December 31, 2023Developed real-time streaming pipelines using Kafka and Spark Streaming to process high-volume retail transaction and inventory datasets with low latency. Built event-driven architectures using Kafka, Spark Streaming, Pub/Sub, and Dataflow. Created ETL workflows for structured, semi-structured, and streaming ingestion into cloud analytics platforms. Implemented real-time inventory tracking and operational monitoring; optimized for low latency, scalability, high availability, and fault tolerance. Managed distributed processing environments including Kafka clusters, Dataproc, and Spark workloads. Built incremental and CDC-based pipelines to improve data freshness, designed dimensional models for forecasting/reporting, and optimized BigQuery performance using partitioning, clustering, and materialized views. Added real-time validation, anomaly detection, and alerting; integrated IAM security, encryption, and secure transfers. Automated deployment/orchestration via CI/CD and infrastructure
Big Data Developer at RBC Bank Group
June 1, 2020 - April 30, 2022Designed cloud-based data lake and warehouse solutions on Azure analytics platforms. Integrated data from banking, transactional, and third-party systems into centralized analytics environments. Built distributed Spark processing jobs for financial reconciliation, compliance reporting, and operational analytics. Optimized Spark workloads and SQL query performance for large-scale datasets. Implemented data cleansing, transformation, enrichment, and validation processes to improve data accuracy. Automated PII masking and de-identification for enterprise security/compliance (PIPEDA, HIPAA). Supported analytics and BI/dashboarding teams with reporting solutions.
Data Analyst at Shopify Inc.
November 1, 2017 - May 31, 2020Analyzed large-scale e-commerce transaction datasets using SQL and Python to generate operational insights, customer analytics, and sales performance reports. Developed ETL pipelines using Python, SQL, and Spark for automated extraction, transformation, and reporting. Created dashboards and visualizations for sales, customer behavior, marketing performance, and operational analytics. Designed and optimized reporting data models (star schema). Performed data cleansing, reconciliation, validation, and QA across enterprise datasets. Supported A/B testing analysis, cohort analysis, customer segmentation, and marketing attribution reporting. Automated recurring reporting workflows using Airflow and SQL scheduling tools; collaborated with product, finance, marketing, and analytics for requirements.
Education
Master of Education (M.Ed.) at Osmania University, Hyderabad, India
January 11, 2030 - August 26, 2026Bachelor of Science (B.Sc.) at St. Francis College for Women, Hyderabad, India
January 11, 2030 - August 26, 2026Qualifications
AWS Certified Developer
January 11, 2030 - August 26, 2026AWS Certified Professional
January 11, 2030 - August 26, 2026Professional Scrum Master
January 11, 2030 - August 26, 2026Industry Experience
Financial Services, Education, Retail, Other
Experience Level
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Intermediate
Intermediate
Hire a Data Engineer
We have the best data engineer experts on Twine. Hire a data engineer in Toronto today.