Hi, I’m Dinan Biswas. I’m a senior software engineer with over 11 years of experience spanning Hadoop ecosystems, data engineering, and cloud technologies. I enjoy building scalable data pipelines, architecting cloud-native solutions, and mentoring teams to deliver high-quality, end-to-end data workloads. Beyond coding, I’m passionate about collaboration, continuous learning, and delivering practical solutions that align with business goals. I thrive in agile environments and love translating complex data problems into robust, maintainable architectures.

Dinan Biswas

Hi, I’m Dinan Biswas. I’m a senior software engineer with over 11 years of experience spanning Hadoop ecosystems, data engineering, and cloud technologies. I enjoy building scalable data pipelines, architecting cloud-native solutions, and mentoring teams to deliver high-quality, end-to-end data workloads. Beyond coding, I’m passionate about collaboration, continuous learning, and delivering practical solutions that align with business goals. I thrive in agile environments and love translating complex data problems into robust, maintainable architectures.

Available to hire

Hi, I’m Dinan Biswas. I’m a senior software engineer with over 11 years of experience spanning Hadoop ecosystems, data engineering, and cloud technologies. I enjoy building scalable data pipelines, architecting cloud-native solutions, and mentoring teams to deliver high-quality, end-to-end data workloads.

Beyond coding, I’m passionate about collaboration, continuous learning, and delivering practical solutions that align with business goals. I thrive in agile environments and love translating complex data problems into robust, maintainable architectures.

See more

Experience Level

Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Intermediate
Intermediate
Intermediate
See more

Language

English
Fluent

Work Experience

Senior Software Engineer / Data Architect at Bank of Montreal
November 1, 2021 - November 22, 2025
Project Details 2: Fraud Detection and Customer Insights - Developed an ETL pipeline in Databricks to extract transactional data, transform it to identify fraudulent activities, derive customer insights, and load into the data warehouse for analysis. Implemented AWS Glue crawlers for automated data source discovery and cataloging; optimized PySpark jobs for performance and scalability; orchestrated loads into Redshift. Led the data analysis in collaboration with business stakeholders to derive actionable insights.
Senior Software Engineer / Data Architect at Bank of Montreal
November 1, 2021 - November 22, 2025
Led the migration of on-premises data applications to AWS, centralizing data into a single Redshift-based repository to enable data sharing across multiple applications. Collaborated with cross-functional teams to assess migration feasibility, dependencies, and risk, and created a comprehensive migration plan with timelines and milestones. Implemented infrastructure as code using AWS CloudFormation and AWS CDK, ensuring repeatable and secure deployments. Worked with security teams to establish best practices for resource hardening and compliance. Designed automated monitoring and logging via CloudWatch and the ELK Stack to improve system visibility and incident response. Mentored team members to enable knowledge transfer through code reviews and hands-on coaching. Project details include Fraud Detection and Customer Insights using Databricks PySpark for ETL, automated data cataloging with Glue crawlers, and loading transformed data into Redshift for BI analysis. Environment: PySpark, P
Technical Lead, App Engineer at Anthem Inc (Legato Health Technologies)
October 1, 2021 - October 1, 2021
ETL Data Lake (IDL): Created a centralized repository for Anthem’s data using AWS, exposing cross-account connectivity via S3, REST API, and streaming. Tokenized, cleaned, and processed data before storing in Hive/Redshift for BI reporting. Automation Framework: Built a framework for sanity checks on data loads across Hive, Redshift, Snowflake, Teradata, Elastic Search, and RDS, enabling cross-database data comparisons (e.g., Teradata-to-Hive). Implemented end-to-end PySpark-based ETL pipelines and shell scripting. Environment: PySpark, Shell Scripting, Elastic Search, Hive, Snowflake, RDS, Redshift, Airflow, SageMaker, AWS Lambda, CloudFormation, API Gateway; CI/CD with Bamboo.
Associate Data Scientist at UnitedHealth Group Optum Global Solution
September 1, 2019 - September 1, 2019
Automated OCR extraction of ICD codes from prescription PDFs (EMR) to support CMS funding submissions. Implemented Hadoop Streaming and PySpark workflows to process documents, with data modeled in HBase and Hive, and integrated Kafka messaging for streaming. Migrated legacy Hadoop Streaming pipelines to PySpark and developed a simple search utility using Whoosh. Leveraged OpenCV-based OCR improvements and collaborated across teams to define data models and testing strategies in a big data ecosystem.
Software Engineer at Atmecs Technologies Pvt. Ltd
April 1, 2017 - April 1, 2017
Migrated a client’s aging Oracle data environment to a scalable cloud-based architecture to enable self-service data mashups across business domains. Led data-mart migration to S3 and Redshift; implemented complex PL/SQL-like transformations using Hadoop Streaming with Pig and Hive; designed and scheduled Oozie workflows; contributed to BI tooling (Birst) for analytics; established ETL governance and deployment automation.

Education

Bachelor of Technology in Computer Science and Engineering at Jawaharlal Nehru Technological University
January 11, 2030 - November 22, 2025

Qualifications

AWS Solutions Architect - Associate
January 11, 2030 - November 22, 2025

Industry Experience

Software & Internet, Financial Services, Healthcare, Professional Services, Education

Experience Level

Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Intermediate
Intermediate
Intermediate
See more