Azure DataBricks data engineer with 8+ years of experience architecting and implementing enterprise-scale batch and real-time data platforms. Proven ability to lead data engineering teams and design scalable Azure DataBricks Medallion Architecture pipelines using PySpark, Spark Structured Streaming, and Delta Lake (DLT/DLT pipelines, Auto Loader, Unity Catalog, and Lakeflow jobs).

PRERNA SHARMA

Azure DataBricks data engineer with 8+ years of experience architecting and implementing enterprise-scale batch and real-time data platforms. Proven ability to lead data engineering teams and design scalable Azure DataBricks Medallion Architecture pipelines using PySpark, Spark Structured Streaming, and Delta Lake (DLT/DLT pipelines, Auto Loader, Unity Catalog, and Lakeflow jobs).

Available to hire

Azure DataBricks data engineer with 8+ years of experience architecting and implementing enterprise-scale batch and real-time data platforms. Proven ability to lead data engineering teams and design scalable Azure DataBricks Medallion Architecture pipelines using PySpark, Spark Structured Streaming, and Delta Lake (DLT/DLT pipelines, Auto Loader, Unity Catalog, and Lakeflow jobs).

See more

Experience Level

Expert
Expert
Expert
Expert
Expert
Intermediate
Intermediate

Work Experience

Lead Azure DataBricks Data Engineer at EPAM Systems
June 1, 2023 - Present
Led enterprise-scale Azure DataBricks implementations for both batch and real-time platforms. Designed and optimized Medallion architecture pipelines using PySpark, Spark Structured Streaming, Delta Lake, Delta Live Tables, and Delta Sharing concepts. Built data ingestion and transformation frameworks using Databricks Auto Loader, file notification modes, schema evolution, and incremental processing. Implemented CDC/SCD Type 2 using Delta Lake Merge, alongside performance optimizations such as Optimize, Vacuum, Z-Ordering, partition tuning, and file compaction. Implemented governance with Unity Catalog and secure data access patterns. Automated deployments using Azure DevOps and Databricks Asset Bundles (DAB), and integrated AI solutions via Azure OpenAI and FastAPI-based services.
Senior Data Engineer at Ticketmaster
October 1, 2022 - June 1, 2023
Worked on Azure DataBricks and Delta Lake pipelines for scalable ingestion and transformation. Delivered CDC and SCD frameworks using Delta Lake Merge operations. Developed and maintained streaming and batch ETL using Spark Structured Streaming and Delta Live Tables. Performed performance tuning and monitoring using Spark UI/executor logs, and contributed to CI/CD pipelines for reliable production releases. Supported production operations, troubleshooting, and ongoing optimization for data workflows.
Azure Data Engineer at Publicis Sapient
April 1, 2021 - October 1, 2022
Built scalable ETL frameworks using PySpark and Spark SQL, and migrated enterprise workloads from on-prem systems to Azure. Developed reusable ingestion frameworks for Terra data and DB2 sources. Implemented CI/CD pipelines and data validation frameworks; created Azure Data Factory and PySpark based enterprise ETL solutions. Built Spark-based transformation frameworks for reporting and analytics and supported production deployments, CI/CD execution, and monitoring.

Education

Bachelor of Engineering (Computer Science & Engineering) at Chitkara Institute of Engineering
January 1, 2018 - April 1, 2021

Qualifications

Add your qualifications or awards here.

Industry Experience

Software & Internet, Professional Services, Financial Services, Media & Entertainment