I'm an AI/ML engineer who loves turning research into scalable production systems. Over the past ~4 years I’ve designed and deployed transformer-based models (LLMs like GPT, BERT, and LLaMA) and built multimodal pipelines for NLP, vision, and recommendations. I enjoy collaborating with cross-functional teams to deliver AI-powered products that are reliable, efficient, and explainable. In my work I emphasize robust MLOps, responsible AI governance, and cloud-native deployments. I optimize performance and latency, implement Retrieval-Augmented Generation (RAG) with vector databases, and craft end-to-end ML workflows—from data preprocessing to monitoring and retraining—on Azure and AWS. I’m passionate about making AI accessible in real-world applications while maintaining safety and transparency.

Subhash Mothukuru

I'm an AI/ML engineer who loves turning research into scalable production systems. Over the past ~4 years I’ve designed and deployed transformer-based models (LLMs like GPT, BERT, and LLaMA) and built multimodal pipelines for NLP, vision, and recommendations. I enjoy collaborating with cross-functional teams to deliver AI-powered products that are reliable, efficient, and explainable. In my work I emphasize robust MLOps, responsible AI governance, and cloud-native deployments. I optimize performance and latency, implement Retrieval-Augmented Generation (RAG) with vector databases, and craft end-to-end ML workflows—from data preprocessing to monitoring and retraining—on Azure and AWS. I’m passionate about making AI accessible in real-world applications while maintaining safety and transparency.

Available to hire

I’m an AI/ML engineer who loves turning research into scalable production systems. Over the past ~4 years I’ve designed and deployed transformer-based models (LLMs like GPT, BERT, and LLaMA) and built multimodal pipelines for NLP, vision, and recommendations. I enjoy collaborating with cross-functional teams to deliver AI-powered products that are reliable, efficient, and explainable.

In my work I emphasize robust MLOps, responsible AI governance, and cloud-native deployments. I optimize performance and latency, implement Retrieval-Augmented Generation (RAG) with vector databases, and craft end-to-end ML workflows—from data preprocessing to monitoring and retraining—on Azure and AWS. I’m passionate about making AI accessible in real-world applications while maintaining safety and transparency.

See more

Experience Level

Expert
Expert
Expert
Expert
Expert
Intermediate

Work Experience

Machine Learning Engineer at Scale AI
June 1, 2025 - Present
Designed, trained, and fine-tuned transformer-based LLMs (GPT, BERT, LLaMA) for NLP tasks; built automated ML pipelines with Python, TensorFlow, PyTorch, and Hugging Face; implemented real-time data ingestion with Spark and Airflow; deployed MLOps workflows on Azure, enabling reproducible training, monitoring, and large-scale production deployment; optimized inference with ONNX, TensorRT, FastAPI; implemented vector databases (Pinecone, FAISS) with Retrieval-Augmented Generation (RAG) to enhance semantic search across 5+ enterprise products.
Software Machine Learning Engineer at Meta
June 1, 2020 - May 1, 2023
Developed and fine-tuned large language models (LLMs) using PyTorch and LLaMA 2 for generative ad copy; built vision diffusion models (Emu, EmuEdit, EmuVideo) for image/background generation and animated creatives; conducted prompt tuning and RLHF experiments; engineered real-time ad recommendation models (GEM) contributing to lifts in conversions; designed scalable inference pipelines with ONNX, TorchScript, and AWS; led canary deployments and A/B testing; built multi-modal pipelines and integrated AI features via RESTful APIs.

Education

Master's in Computer Science and Software Engineering at Concordia University Wisconsin
January 11, 2030 - June 29, 2026

Qualifications

Introduction to GenerativeAI
January 11, 2030 - June 29, 2026
Microsoft Certified: Azure AI Fundamentals
January 11, 2030 - June 29, 2026
LangChain for LLM Application Development
January 11, 2030 - June 29, 2026
Building Ambient Agents with LangGraph
January 11, 2030 - June 29, 2026
Fine-tuning Language Models with Azure AI Foundry
January 11, 2030 - June 29, 2026
AI Agents in LangGraph
January 11, 2030 - June 29, 2026
Building AI Voice Agents for Production
January 11, 2030 - June 29, 2026

Industry Experience

Software & Internet, Media & Entertainment