Available to hire
Data/AI Engineer with 3+ years of experience building scalable ETL pipelines, AI-powered applications, and cloud-native analytics solutions. Skilled in Python, SQL, PySpark, Azure Data Factory, machine learning, and Retrieval-Augmented Generation (RAG) systems.
Experienced in production-grade AI services and large-scale data workflow optimization across Azure and AWS, with deployments using Docker and MLflow. Passionate about delivering business-driven insights through reliable, scalable data and AI platforms.
Skills
Experience Level
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Intermediate
Intermediate
Intermediate
Intermediate
Beginner
Beginner
Language
English
Advanced
Work Experience
Data/AI Engineer at UGenome AI
August 1, 2025 - PresentDeveloped scalable ETL pipelines using Azure Data Factory, PySpark, and SQL to automate healthcare data ingestion workflows, processing over 5M+ records monthly and improving operational efficiency by 40%. Architected and deployed HIPAA-compliant RAG-based AI systems using LangChain and embedding pipelines for semantic search and intelligent document retrieval. Built production-grade AI services using FastAPI, Docker, and MLflow with automated CI/CD deployment, reducing model deployment time by 35% across Azure and AWS. Implemented data validation and transformation frameworks for schema enforcement, anomaly detection, missing-value handling, and data quality optimization. Enhanced LLM response relevance via embedding chunking, vector search optimization, and prompt engineering. Integrated SQL-based reporting pipelines with Power BI dashboards and collaborated with cross-functional teams in Agile sprints to deliver scalable AI-driven analytics solutions.
Data Scientist at Persistent Systems
February 1, 2021 - July 1, 2023Analyzed and processed 3M+ industrial IoT and sensor data records using Python, SQL, and statistical analysis to support predictive analytics initiatives generating over $500K in annual operational savings. Developed and optimized predictive maintenance and forecasting models using Scikit-learn and XGBoost, achieving 96% accuracy. Built anomaly detection and time-series forecasting systems to identify early equipment fault patterns, reducing downtime by 30%. Automated enterprise ETL workflows, data transformation pipelines, and reporting using Python, SQL, PySpark, and workflow automation tools, reducing manual reporting by 40%. Designed relational data models and enterprise data warehousing solutions using SQL Server and MySQL to support KPI tracking and advanced analytics. Built real-time monitoring dashboards using Flask, Plotly, Power BI, and Tableau, and worked in Agile environments to deliver production-ready ML solutions and reporting applications.
Education
Masters of Science in Data Science at University of Arizona
August 1, 2023 - May 1, 2025Masters of Science in Data Science at University of Arizona
August 1, 2023 - May 1, 2025Masters of Science in Data Science at University of Arizona
August 1, 2023 - May 31, 2025Qualifications
Industry Experience
Healthcare, Telecommunications, Software & Internet, Computers & Electronics, Professional Services, Manufacturing
Skills
Experience Level
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Intermediate
Intermediate
Intermediate
Intermediate
Beginner
Beginner
Hire a Data Scientist
We have the best data scientist experts on Twine. Hire a data scientist in Tucson today.