Hi, I’m Divya Babulal Shah. I’m a Full Stack AI Engineer with 3+ years of delivering LLM-powered applications, data pipelines, agentic workflows, and customer-facing AI tools from prototype to production using Python, TypeScript, React, and Next.js. I’ve consistently framed problems with business owners, shipped working prototypes quickly, and presented results clearly to non-technical stakeholders. I have a track record of eliminating manual workflow effort, cutting processing time by 70%, and deploying tools that users can feel and trust. I enjoy collaborating with cross-functional teams, designing end-to-end solutions, and turning complex data challenges into actionable insights. I’ve built scalable ML/Ops pipelines, improved model recall, and delivered dashboards and interfaces that empower non-technical users to interact with advanced AI capabilities in plain English.

Divya Babulal Shah

Hi, I’m Divya Babulal Shah. I’m a Full Stack AI Engineer with 3+ years of delivering LLM-powered applications, data pipelines, agentic workflows, and customer-facing AI tools from prototype to production using Python, TypeScript, React, and Next.js. I’ve consistently framed problems with business owners, shipped working prototypes quickly, and presented results clearly to non-technical stakeholders. I have a track record of eliminating manual workflow effort, cutting processing time by 70%, and deploying tools that users can feel and trust. I enjoy collaborating with cross-functional teams, designing end-to-end solutions, and turning complex data challenges into actionable insights. I’ve built scalable ML/Ops pipelines, improved model recall, and delivered dashboards and interfaces that empower non-technical users to interact with advanced AI capabilities in plain English.

Available to hire

Hi, I’m Divya Babulal Shah. I’m a Full Stack AI Engineer with 3+ years of delivering LLM-powered applications, data pipelines, agentic workflows, and customer-facing AI tools from prototype to production using Python, TypeScript, React, and Next.js. I’ve consistently framed problems with business owners, shipped working prototypes quickly, and presented results clearly to non-technical stakeholders. I have a track record of eliminating manual workflow effort, cutting processing time by 70%, and deploying tools that users can feel and trust.

I enjoy collaborating with cross-functional teams, designing end-to-end solutions, and turning complex data challenges into actionable insights. I’ve built scalable ML/Ops pipelines, improved model recall, and delivered dashboards and interfaces that empower non-technical users to interact with advanced AI capabilities in plain English.

See more

Experience Level

Expert
Expert
Expert
Expert
Expert
Intermediate

Language

English
Fluent

Work Experience

Data Engineer at Community Dreams Foundation
October 1, 2025 - Present
Built an end-to-end AI data extraction pipeline automating collection of clean energy and housing incentive programs across 25+ sources using Playwright, LangChain (Gemini), and rule-based validation, reducing manual research from days to minutes and delivering a deduplicated dataset across 5 incentive categories. Reduced manual document processing time by 70% by building and shipping a multimodal RAG pipeline as a FastAPI microservice using LangChain and ChromaDB, enabling non-technical stakeholders to query complex documents in plain English via a Streamlit interface. Delivered an end-to-end telecom churn prediction system, improving recall by 13% using XGBoost with Optuna tuning and MLflow tracking, and deployed it via an MLOps pipeline (FastAPI, Docker, AWS ECS, GitHub Actions), reducing deployment cycles by 80%.
Software Data Engineer Co-op at Fidelity Investments
July 1, 2024 - December 1, 2024
Delivered 2M+ weekly records to investment analytics teams by building Python ETL pipelines integrating 30+ external data sources into Snowflake via AWS S3 and Snowpipe, eliminating 6 hours of manual weekly ingestion with fully automated daily refresh cycles. Reduced manual reporting effort by 80% by partnering with 20+ investment analysts to translate business requirements into automated Tableau dashboards tracking workforce trends across 500+ public companies. Maintained 99.9% pipeline reliability across 2M+ weekly records by implementing a Python/SQL data quality framework with row-count reconciliation and anomaly detection in Snowflake, catching 15+ critical data issues weekly before downstream consumption.
Data Engineer at Tata Consultancy Services (TCS)
June 1, 2021 - November 1, 2022
Delivered 20+ custom eForms for public sector and EdTech clients using HTML, CSS, JavaScript, and jQuery with client-side validation and dynamic field logic, reducing data entry errors by 35% across deployed production forms. Identified $20K+ in revenue leakage and improved data quality by 25% by building SQL-based profiling scripts for payment anomaly detection, presenting findings and recommendations directly to business stakeholders. Accelerated data issue resolution by 30% by building Python and SQL automation scripts for root cause analysis across 50+ production forms, and enhanced operational visibility by creating Power BI dashboards tracking 1,500+ daily submissions for client leadership teams.

Education

Master of Science in Data Analytics Engineering at Northeastern University
September 1, 2023 - August 1, 2025
Bachelor of Technology in Electrical Engineering at University of Mumbai
August 1, 2017 - June 1, 2021

Qualifications

AWS Certified Data Engineer - Associate
December 1, 2024 - June 29, 2026
Databricks Certified Data Engineer Associate
May 1, 2026 - June 29, 2026
Agentic AI - DeepLearning.AI
November 1, 2025 - June 29, 2026
Deep Learning Specialization - Coursera
July 1, 2020 - June 29, 2026

Industry Experience

Software & Internet, Financial Services, Education, Professional Services, Energy & Utilities, Telecommunications

Experience Level

Expert
Expert
Expert
Expert
Expert
Intermediate