Hello! I'm Vinay Kumar Reddy Veereddy, an AI/ML engineer with 4+ years of experience delivering end-to-end machine learning and Generative AI solutions across autonomous driving and fintech domains. I thrive on turning complex data into production-grade systems, from data curation and model training to scalable deployment and governance. I specialize in building production-ready computer vision, NLP, and LLM-powered solutions using AWS, GCP, and scalable MLOps architectures. I’ve led full lifecycle perception stacks, document intelligence pipelines with Retrieval-Augmented Generation, and continuous improvement through measurable impact, such as faster retraining cycles and automated processes.

Vinay Kumar Reddy Veereddy

Hello! I'm Vinay Kumar Reddy Veereddy, an AI/ML engineer with 4+ years of experience delivering end-to-end machine learning and Generative AI solutions across autonomous driving and fintech domains. I thrive on turning complex data into production-grade systems, from data curation and model training to scalable deployment and governance. I specialize in building production-ready computer vision, NLP, and LLM-powered solutions using AWS, GCP, and scalable MLOps architectures. I’ve led full lifecycle perception stacks, document intelligence pipelines with Retrieval-Augmented Generation, and continuous improvement through measurable impact, such as faster retraining cycles and automated processes.

Available to hire

Hello! I’m Vinay Kumar Reddy Veereddy, an AI/ML engineer with 4+ years of experience delivering end-to-end machine learning and Generative AI solutions across autonomous driving and fintech domains. I thrive on turning complex data into production-grade systems, from data curation and model training to scalable deployment and governance.

I specialize in building production-ready computer vision, NLP, and LLM-powered solutions using AWS, GCP, and scalable MLOps architectures. I’ve led full lifecycle perception stacks, document intelligence pipelines with Retrieval-Augmented Generation, and continuous improvement through measurable impact, such as faster retraining cycles and automated processes.

See more

Experience Level

Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Intermediate
Intermediate
Intermediate
Intermediate
See more

Language

English
Fluent

Work Experience

AI/ML Engineer at General Motors
March 1, 2024 - Present
Led end-to-end development of perception module for ADAS/autonomous driving pipeline using YOLOv8 and TensorFlow, improving object detection latency by 22% across real-time driving scenarios. Owned full lifecycle from data curation and model training to edge deployment on Kubernetes-based inference. Built Retrieval-Augmented Generation pipelines by integrating vector databases and LangChain for semantic retrieval. Drove 30% faster retraining cycles and 90% process automation. Implemented scalable ML pipelines on AWS SageMaker and GitHub Actions, orchestrated in production with model monitoring and alerting.
AI/ML Engineer
March 1, 2024 - Present
Led end-to-end development of perception modules for ADAS in autonomous driving, including dataset curation, model training, and edge deployment. Improved inference latency and accuracy with YOLOv8 and TensorFlow, deploying on Kubernetes-based inference infrastructure. Architected distributed ML pipelines and Retrieval-Augmented Generation (RAG) using vector stores and LangChain to integrate with vision and sensor data. Implemented model governance, monitoring, and proactive retraining to sustain production performance.
AI/ML Engineer
December 1, 2020 - August 1, 2023
Led development of AI-powered document intelligence platform using BERT and Hugging Face models, automating extraction of structured data from financial records. Designed Retrieval-Augmented Generation (RAG) pipelines using LangChain and Pinecone vector DB, enabling contextual QA over financial knowledge bases. Improved information retrieval accuracy by 35% for customer support and internal analytics. Built and fine-tuned transformer models for multi-label classification of service requests, improving routing accuracy by 18%. Deployed ML models as REST APIs on GCP Vertex AI for real-time inference on web and mobile apps. Implemented CI/CD with Jenkins and GitHub Actions, reducing release cycle time by 40%. Built a semantic search engine using embeddings and vector similarity for document discovery, boosting user search relevance by 28%.
Software Engineer
December 1, 2020 - August 1, 2023
Led the development of an AI-powered document intelligence platform using BERT and HuggingFace. Designed Retrieval-Augmented Generation (RAG) pipelines with LangChain and vector stores to answer fintech knowledge queries. Built scalable data engineering pipelines with PySpark, Kafka, and DataBricks; deployed ML models as REST APIs on GCP Vertex AI with sub-200ms latency. Implemented CI/CD with Jenkins and GitHub Actions; built a semantic search engine with embeddings for document discovery and knowledge retrieval; created dashboards to monitor model performance.

Education

Master of Science in Computer Science at University of Central Missouri
January 11, 2030 - June 29, 2026
Bachelor of Engineering in Electronics and Communication at Jawaharlal Nehru Technological University, India
January 11, 2030 - June 29, 2026
Master of Science in Computer Science at University of Central Missouri
January 11, 2030 - June 29, 2026
Bachelor of Engineering in Electronics and Communication at Jawaharlal Nehru Technological University, India
January 11, 2030 - June 29, 2026

Qualifications

AWS Certified Machine Learning – Specialty
January 11, 2030 - June 29, 2026
AWS Certified Machine Learning – Specialty
January 11, 2030 - June 29, 2026

Industry Experience

Software & Internet, Media & Entertainment, Professional Services, Transportation & Logistics, Other, Financial Services