Hello, I'm Gopu Sagar Reddy, a Senior Data Engineer with over a decade of experience delivering enterprise-scale data platforms across financial services, healthcare, telecom, and e-commerce. I design and implement scalable ETL/ELT pipelines, streaming architectures, and cloud-native data warehouses on AWS, Azure, and GCP, with a strong emphasis on governance, regulatory compliance, and observability. I enjoy collaborating with business, risk, and compliance stakeholders to translate complex data into impactful data products, regulatory reports, and executive dashboards. My hands-on strengths include Spark, PySpark, Kafka, Airflow, Snowflake, and modern data tooling, along with experience in GenAI-enabled analytics and multi-cloud architectures.

Gopu Sagar Reddy

Hello, I'm Gopu Sagar Reddy, a Senior Data Engineer with over a decade of experience delivering enterprise-scale data platforms across financial services, healthcare, telecom, and e-commerce. I design and implement scalable ETL/ELT pipelines, streaming architectures, and cloud-native data warehouses on AWS, Azure, and GCP, with a strong emphasis on governance, regulatory compliance, and observability. I enjoy collaborating with business, risk, and compliance stakeholders to translate complex data into impactful data products, regulatory reports, and executive dashboards. My hands-on strengths include Spark, PySpark, Kafka, Airflow, Snowflake, and modern data tooling, along with experience in GenAI-enabled analytics and multi-cloud architectures.

Available to hire

Hello, I’m Gopu Sagar Reddy, a Senior Data Engineer with over a decade of experience delivering enterprise-scale data platforms across financial services, healthcare, telecom, and e-commerce. I design and implement scalable ETL/ELT pipelines, streaming architectures, and cloud-native data warehouses on AWS, Azure, and GCP, with a strong emphasis on governance, regulatory compliance, and observability.

I enjoy collaborating with business, risk, and compliance stakeholders to translate complex data into impactful data products, regulatory reports, and executive dashboards. My hands-on strengths include Spark, PySpark, Kafka, Airflow, Snowflake, and modern data tooling, along with experience in GenAI-enabled analytics and multi-cloud architectures.

See more

Experience Level

Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
See more

Language

English
Fluent

Work Experience

Sr. Data Engineer at JP Morgan Chase
January 1, 2023 - November 4, 2025
Architected and maintained end-to-end ETL/ELT pipelines using Airflow, Python, and BigQuery to automate ingestion and transformation of complex regulatory and risk datasets. Leveraged Palantir Foundry to integrate financial and insurance-regulatory data into governed ontologies, and used Foundry Code Workbooks and Workflow Modules to orchestrate ELT pipelines with PySpark and SQL. Partnered with data scientists and risk experts to ensure data accuracy, lineage, and compliance for regulatory reporting and risk modeling. Integrated Snowflake with key data sources, optimizing queries, partitions, and materialized views for enterprise analytics. Built live dashboards in Tableau/Power BI connected to Snowflake to monitor credit risk exposure and compliance metrics. Implemented data governance workflows (lineage, quality checks, audit trails) and automated reconciliation/exceptions reporting; designed pipelines to support GenAI anomaly detection and regulatory monitoring.
Sr. Data Engineer at Google
December 1, 2022 - December 1, 2022
Designed and deployed real-time data pipelines in GCP using BigQuery, Dataflow, Pub/Sub, and Cloud Storage to support regulatory and analytics workloads. Partnered with Data Science teams to prepare ML-ready datasets for GenAI models and accelerated experimentation. Optimized BigQuery workloads with clustering, partitioning, and caching for terabyte-scale analytics and built Looker semantic models for global advertising metrics. Automated clickstream ingestion pipelines with Python and Cloud Functions; established pipeline observability with Cloud Logging and Cloud Monitoring. Defined enterprise schema naming conventions, created technical runbooks, and documented data lineage. Delivered experimentation results from A/B tests and provided runbooks for production readiness; participated in peer code reviews to maintain security and engineering standards.
Sr. Data Engineer at AT&T
February 1, 2021 - February 1, 2021
Built Azure Data Factory pipelines integrating Teradata, Hadoop, and SQL Server into enterprise data lakes. Developed PySpark and Hive scripts for billing, call records, and usage analytics; automated Tableau/Power BI dashboards for network performance, churn insights, and customer analytics. Migrated legacy ETL to Hadoop + Spark to improve scalability. Designed metadata-driven pipelines with monitoring and scheduling frameworks; partnered with business teams to create customer segmentation models for churn prevention. Wrote reusable SQL scripts and stored procedures for billing analytics and governance of data assets. Worked in Agile environments with Jira and Rally; supported ad-hoc reporting and data quality improvements.
Sr. Data Engineer at CVS Health
November 1, 2018 - November 1, 2018
Extracted and transformed claims and pharmacy data from Oracle/Teradata into enterprise warehouses; built SQL workflows and Python automation scripts for fraud detection and compliance reporting. Built Palantir Foundry pipelines and Python code to process healthcare and insurance claim data, delivering CMS and HIPAA-regulated reports with audit trails and data validation checks. Implemented governance practices, dashboards (Tableau/Power BI), and data quality controls; migrated legacy SSRS/Excel reporting to modern BI platforms. Collaborated with business users to define KPIs for claims, member behavior, and compliance, and delivered automated recurring extracts and regulatory outputs.
Sr. Data Engineer at Flipkart
November 1, 2015 - November 1, 2015
Developed Hive/SQL/Python ETL workflows for sales, supply chain, and campaign analytics; built Tableau dashboards for customer engagement, vendor performance, and pricing optimization. Automated data quality checks and supported Hadoop migration for large-scale e-commerce workloads. Partnered with marketing to deliver customer segmentation and ROI analyses; designed metadata-driven pipelines and documented data mappings for governance. Assisted in exploratory analysis for new product launches and campaign effectiveness; produced KPI dashboards and reports to support business decisions.

Education

Add your educational history here.

Qualifications

Add your qualifications or awards here.

Industry Experience

Financial Services, Healthcare, Telecommunications, Software & Internet