Hi, I'm Bala Raviteja Karna, an AI/ML Engineer with over 4 years of experience building, deploying, and optimizing ML, NLP, and Generative AI solutions in production. I have hands-on experience delivering end-to-end AI systems, from data engineering and model development to deployment, monitoring, and continuous optimization. I enjoy turning complex business problems into scalable AI systems and collaborating with cross-functional teams to drive data-driven decision making. Whether it's building LLM-powered applications, RAG systems, semantic search platforms, or MLOps pipelines, I focus on reliability, scalability, and measurable impact. I have hands-on experience with Python, PyTorch, TensorFlow, cloud platforms, LangChain, LangGraph, and Hugging Face, and I strive to deliver solutions that improve relevance, reduce latency, and enable informed business decisions.

Bala Raviteja Karna

Hi, I'm Bala Raviteja Karna, an AI/ML Engineer with over 4 years of experience building, deploying, and optimizing ML, NLP, and Generative AI solutions in production. I have hands-on experience delivering end-to-end AI systems, from data engineering and model development to deployment, monitoring, and continuous optimization. I enjoy turning complex business problems into scalable AI systems and collaborating with cross-functional teams to drive data-driven decision making. Whether it's building LLM-powered applications, RAG systems, semantic search platforms, or MLOps pipelines, I focus on reliability, scalability, and measurable impact. I have hands-on experience with Python, PyTorch, TensorFlow, cloud platforms, LangChain, LangGraph, and Hugging Face, and I strive to deliver solutions that improve relevance, reduce latency, and enable informed business decisions.

Available to hire

Hi, I’m Bala Raviteja Karna, an AI/ML Engineer with over 4 years of experience building, deploying, and optimizing ML, NLP, and Generative AI solutions in production. I have hands-on experience delivering end-to-end AI systems, from data engineering and model development to deployment, monitoring, and continuous optimization. I enjoy turning complex business problems into scalable AI systems and collaborating with cross-functional teams to drive data-driven decision making.

Whether it’s building LLM-powered applications, RAG systems, semantic search platforms, or MLOps pipelines, I focus on reliability, scalability, and measurable impact. I have hands-on experience with Python, PyTorch, TensorFlow, cloud platforms, LangChain, LangGraph, and Hugging Face, and I strive to deliver solutions that improve relevance, reduce latency, and enable informed business decisions.

See more

Experience Level

Expert
Expert
Expert
Expert
Expert
Expert
Expert
Intermediate
See more

Language

English
Fluent

Work Experience

AI/ML Engineer at OpenAI
February 1, 2025 - Present
Developed enterprise-grade Generative AI applications leveraging OpenAI GPT models, LangChain, and Retrieval-Augmented Generation (RAG). Designed and deployed scalable RAG pipelines with vector search, embeddings, and LLM orchestration; improved answer relevance by 35%. Fine-tuned transformer-based models using LoRA and PEFT for domain-specific NLP. Built scalable ML workflows with PyTorch, TensorFlow, and Apache Airflow; automated data ingestion, feature engineering, model training, and deployment. Reduced deployment timelines by 40% via AWS SageMaker deployment pipelines, Docker, and CI/CD. Established MLOps/LLMOps with MLflow, Weights & Biases, Jenkins, and Git-based deployment. Developed monitoring dashboards (Grafana, Plotly) to track performance and drift. Implemented FastAPI inference services and containerized deployment workflows for scalability. Led end-to-end delivery from requirements to deployment and monitoring, collaborating with product, software, and business stakehold
ML Engineer at Amazon
June 1, 2020 - July 1, 2023
Developed and deployed ML models using Scikit-Learn, XGBoost, and Random Forest to improve forecasting and recommendations. Built NLP solutions with BERT, spaCy, and NLTK for sentiment analysis and intent classification. Improved demand forecasting accuracy by 25% with LSTM and Prophet-based time-series models. Designed and optimized PySpark, Pandas, and Apache Airflow pipelines for large-scale analytics. Developed semantic search solutions using transformer models and embeddings to enhance information retrieval. Collaborated with software and data engineering teams to deploy production ML services using Docker, Jenkins, and CI/CD. Leveraged AWS SageMaker, EC2, and S3 for scalable training and lifecycle management. Drove data-driven decision making via actionable analytics.

Education

Master of Science in Cybersecurity at Webster University, San Antonio, TX
August 1, 2023 - May 1, 2025

Qualifications

Add your qualifications or awards here.

Industry Experience

Software & Internet, Professional Services, Media & Entertainment