Data Engineer with an MSc in Data Science & AI and production experience at Amadeus, designing and maintaining large-scale distributed data pipelines on AWS. I specialize in Spark-based ETL systems, Scala/PySpark migration, and improving pipeline performance and stability while processing millions of records in production. I work across data modeling, batch processing, Linux, and distributed architectures, with strong foundations in SQL and automation. I’ve also built CI/CD deployment workflows and developed end-to-end data/ML and data-engineering projects, from ingestion to evaluation.

Aritra Mondal

Data Engineer with an MSc in Data Science & AI and production experience at Amadeus, designing and maintaining large-scale distributed data pipelines on AWS. I specialize in Spark-based ETL systems, Scala/PySpark migration, and improving pipeline performance and stability while processing millions of records in production. I work across data modeling, batch processing, Linux, and distributed architectures, with strong foundations in SQL and automation. I’ve also built CI/CD deployment workflows and developed end-to-end data/ML and data-engineering projects, from ingestion to evaluation.

Available to hire

Data Engineer with an MSc in Data Science & AI and production experience at Amadeus, designing and maintaining large-scale distributed data pipelines on AWS. I specialize in Spark-based ETL systems, Scala/PySpark migration, and improving pipeline performance and stability while processing millions of records in production.

I work across data modeling, batch processing, Linux, and distributed architectures, with strong foundations in SQL and automation. I’ve also built CI/CD deployment workflows and developed end-to-end data/ML and data-engineering projects, from ingestion to evaluation.

See more

Experience Level

Expert
Expert
Expert
Expert
Expert
Expert
Expert
Intermediate
Beginner
Beginner
See more

Language

Hindi
Fluent
Bengali
Fluent
English
Fluent
French
Intermediate
Spanish; Castilian
Beginner
Assamese
Intermediate
Oriya
Advanced

Work Experience

Data Engineer (Apprenticeship) at Amadeus
August 1, 2023 - September 1, 2024
Designed and maintained large-scale ETL pipelines on AWS (EMR, S3) supporting distributed data processing workflows across multiple data sources. Migrated PySpark-based pipelines to Scala, improving processing throughput by 35% and reducing cluster resource consumption via optimized execution. Optimized data ingestion and transformation workflows handling millions of records in production, improving stability and performance. Automated CI/CD deployment workflows using Jenkins and Bash to ensure reliable and repeatable data pipeline delivery.
Android Developer (Intern) at Jungleworks
November 1, 2020 - April 1, 2021
Developed Android applications and integrated backend APIs to support data exchange and application logic. Worked with REST-based data exchange systems, including debugging, validation, and integration testing across client-server workflows.

Education

MSc Data Science & IA at DSTI, France
January 1, 2022 - January 1, 2024
BSc (Hons) Computer Security at Northumbria University, UK
January 1, 2017 - September 1, 2020

Qualifications

Neo4j Certified Professional
January 1, 2024 - August 5, 2026
AWS Solutions Architect Associate
January 11, 2030 - August 5, 2026

Industry Experience

Software & Internet, Professional Services, Computers & Electronics, Financial Services

Experience Level

Expert
Expert
Expert
Expert
Expert
Expert
Expert
Intermediate
Beginner
Beginner
See more