AI Platform/MLOps engineer with 10+ years of experience building and supporting scalable AI and cloud-native platforms across AWS, Azure, and GCP. Strong end-to-end expertise in MLOps and LLMOps, including training, deployment, monitoring, governance, and lifecycle automation for production AI systems.

VAISHAVI FALGUN

AI Platform/MLOps engineer with 10+ years of experience building and supporting scalable AI and cloud-native platforms across AWS, Azure, and GCP. Strong end-to-end expertise in MLOps and LLMOps, including training, deployment, monitoring, governance, and lifecycle automation for production AI systems.

Available to hire

AI Platform/MLOps engineer with 10+ years of experience building and supporting scalable AI and cloud-native platforms across AWS, Azure, and GCP. Strong end-to-end expertise in MLOps and LLMOps, including training, deployment, monitoring, governance, and lifecycle automation for production AI systems.

See more

Experience Level

Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Intermediate
Intermediate
See more

Language

Work Experience

Senior AI Platform Engineer / Lead MLOps Engineer at PayPal
December 1, 2025 - Present
Designed and maintained enterprise AI platforms supporting ML, deep learning, and Generative AI workloads. Architected scalable infrastructure enabling rapid model deployment and operationalization, including reusable platform components and self-service deployment capabilities. Built end-to-end MLOps pipelines for data ingestion, model training, validation, deployment, and monitoring. Managed multi-region GKE clusters and implemented centralized logging and monitoring using Google Cloud Operations Suite, Prometheus, and Grafana. Operationalized MLflow experiment tracking, model registry, and model governance with automated model promotion and retraining workflows. Implemented RAG solutions using LangChain and LangGraph with vector databases, fine-tuned and evaluated open-source LLMs, and added production monitoring including drift detection and performance observability. Provided incident response, capacity planning, and platform optimization, leveraging AI-assisted development tool
DevOps Engineer / ML Platform Engineer at TCS
March 1, 2020 - December 1, 2025
Designed and implemented scalable MLOps platforms supporting NLP and ML workloads. Built reusable deployment frameworks and GitLab CI/CD pipelines for packaging, validation, and deployment. Implemented Infrastructure as Code for repeatable environments and managed model lifecycle using MLflow and AWS SageMaker. Developed NLP solutions using BERT, DistilBERT, Hugging Face, spaCy, and NLTK, including sentiment analysis, text classification, and chatbot solutions, with fine-tuning for domain-specific use cases. Implemented automated data-quality validation and schema verification, containerized workloads with Docker and Kubernetes, and delivered monitoring/drift detection and observability dashboards using Grafana, Kibana/ELK, and Splunk. Collaborated with data science teams to productionize models and optimized cloud infrastructure for cost and performance.
DevOps Engineer / Site Reliability Engineer (SRE) at Virgin Media O2
March 1, 2019 - March 1, 2020
Managed Kubernetes and GKE clusters supporting business-critical applications. Designed CI/CD pipelines for multiple engineering squads. Automated infrastructure provisioning using Terraform and Ansible. Implemented monitoring and alerting with Prometheus, Grafana, and Dynatrace, and improved platform reliability through capacity planning and incident management. Applied AWS Well-Architected Framework best practices in operations and reliability improvements.
DevOps Engineer / Site Reliability Engineer (SRE) at Tech Mahindra
January 1, 2018 - March 1, 2019
Managed AWS infrastructure across production and non-production environments. Automated deployments using Jenkins, GitHub, and Infrastructure as Code. Built Docker images and Kubernetes deployment pipelines, implemented monitoring, logging, and cloud security controls, and supported high-availability architectures and release automation. Designed and deployed cloud architectures across IaaS/PaaS/SaaS and performed AWS provisioning using core services such as EC2, S3, ELB, RDS, IAM, Route 53, VPC/subnets, routing, auto-scaling, CloudFront, CloudWatch, and CloudFormation. Used Terraform and Ansible for infrastructure automation.
Test Automation Analyst at Sai Tek Limited
June 1, 2012 - January 1, 2018
Performed test analysis of user stories, identified requirements to be tested, and participated in Agile ceremonies (standup, sprint planning, retrospectives). Authored feature files using Cucumber based on acceptance criteria and executed tests each sprint. Developed automated tests using Cypress (JavaScript) and Pytest scripts for Selenium automation (Python). Executed and maintained automation scripts, debugging issues and tracking defects in Jira/X-ray. Performed E2E testing and mobile manual testing, and built scripts for build/deployment/maintenance using Jenkins, Docker, Maven, and Bash. Created CI build/deployment infrastructure for multiple projects. Supported UAT and API testing with Postman, performed ETL/data testing, and ran performance/load testing using JMeter.

Education

Add your educational history here.

Qualifications

MSc Computer Science
January 1, 2010 - January 1, 2010
Bachelor of Computer Science
January 1, 2008 - January 1, 2008
MSc Computer Science
January 1, 2010 - August 3, 2026
Bachelor of Computer Science
January 1, 2008 - August 3, 2026
AWS Certified Solutions Architect
January 1, 2020 - August 3, 2026
ISTQB Certified Tester
January 1, 2015 - August 3, 2026

Industry Experience

Software & Internet, Computers & Electronics, Financial Services, Professional Services, Other