I'm Sai Chandra Palle, a Machine Learning Engineer with 5+ years of experience building scalable, production-grade AI systems, including LLM-powered recommendations and multilingual NLP platforms. I excel at end-to-end ML and GenAI—designing data pipelines, training models, and deploying real-time inference on AWS—delivering measurable business impact while improving retrieval diversity, engagement, and latency through generative AI, transformer architectures, and semantic search.

Sai Chandra Palle

I'm Sai Chandra Palle, a Machine Learning Engineer with 5+ years of experience building scalable, production-grade AI systems, including LLM-powered recommendations and multilingual NLP platforms. I excel at end-to-end ML and GenAI—designing data pipelines, training models, and deploying real-time inference on AWS—delivering measurable business impact while improving retrieval diversity, engagement, and latency through generative AI, transformer architectures, and semantic search.

Available to hire

I’m Sai Chandra Palle, a Machine Learning Engineer with 5+ years of experience building scalable, production-grade AI systems, including LLM-powered recommendations and multilingual NLP platforms.

I excel at end-to-end ML and GenAI—designing data pipelines, training models, and deploying real-time inference on AWS—delivering measurable business impact while improving retrieval diversity, engagement, and latency through generative AI, transformer architectures, and semantic search.

See more

Experience Level

Expert
Expert
Expert
Expert
Expert

Language

English
Fluent

Work Experience

Machine Learning Engineer at Pinterest
August 1, 2024 - Present
Led LLM-powered generative retrieval and semantic ranking initiatives to boost diversity and engagement; implemented dense embedding pipelines for semantic search and RAG-based retrieval; developed ANN-based candidate generation with Faiss (IVF-HNSW) for scalable, high-quality recommendations; designed distributed data pipelines on AWS EMR with Spark; built scalable multi-stage inference workflows and microservice-based ML architectures (TorchScript deployment) to improve throughput and latency; leveraged AWS S3 and Iceberg on S3 for feature storage and experiment tracking.
Machine Learning Engineer at Amazon India
January 1, 2020 - July 1, 2023
Enhanced multilingual query understanding by 34% for code-mixed inputs; reduced API latency by 29% via optimized real-time inference pipelines and caching; increased data processing efficiency by 38% with scalable ETL on AWS Glue and S3 for multilingual datasets. Built and fine-tuned transformer-based models (XLM-R, mT5) for multilingual NLP tasks, deployed on AWS (EC2, SageMaker) with monitoring (CloudWatch) and robust backend services.

Education

Master of Science in Computer Science at Southern University and A&M
January 11, 2030 - June 29, 2026
Bachelor of Technology in Computer Science and Engineering at Mahatma Gandhi Institute of Technology
January 11, 2030 - June 29, 2026

Qualifications

Oracle Cloud Infrastructure Generative AI Professional — Focused on LLMs, embeddings, and enterprise-scale Generative AI applications
January 11, 2030 - June 29, 2026
Oracle Cloud Infrastructure Generative AI Professional — Covered cloud-based AI services, model deployment, and scalable AI system design
January 11, 2030 - June 29, 2026

Industry Experience

Software & Internet, Media & Entertainment, Professional Services