I am an AI Software Engineer with 4 years of experience designing, developing, and deploying scalable AI-driven applications, distributed systems, and cloud-native microservices across production environments. I specialize in Generative AI, large language models (LLMs), Retrieval-Augmented Generation (RAG), AI agents, computer vision, and real-time inference optimization using Python, PyTorch, TensorFlow, LangChain, and FastAPI. I have built high-performance data pipelines, vector search systems, and event-driven architectures with Apache Kafka, Redis, Elasticsearch, and Kubernetes. I bring a strong background in MLOps and LLMOps, CI/CD automation, model deployment, observability, and multi-cloud platforms (AWS, Azure, GCP). I thrive on cross-functional collaboration to improve system scalability, reduce inference latency, and deliver reliable production AI solutions.

Devi Nikhitha Mamillapalli

I am an AI Software Engineer with 4 years of experience designing, developing, and deploying scalable AI-driven applications, distributed systems, and cloud-native microservices across production environments. I specialize in Generative AI, large language models (LLMs), Retrieval-Augmented Generation (RAG), AI agents, computer vision, and real-time inference optimization using Python, PyTorch, TensorFlow, LangChain, and FastAPI. I have built high-performance data pipelines, vector search systems, and event-driven architectures with Apache Kafka, Redis, Elasticsearch, and Kubernetes. I bring a strong background in MLOps and LLMOps, CI/CD automation, model deployment, observability, and multi-cloud platforms (AWS, Azure, GCP). I thrive on cross-functional collaboration to improve system scalability, reduce inference latency, and deliver reliable production AI solutions.

Available to hire

I am an AI Software Engineer with 4 years of experience designing, developing, and deploying scalable AI-driven applications, distributed systems, and cloud-native microservices across production environments. I specialize in Generative AI, large language models (LLMs), Retrieval-Augmented Generation (RAG), AI agents, computer vision, and real-time inference optimization using Python, PyTorch, TensorFlow, LangChain, and FastAPI.

I have built high-performance data pipelines, vector search systems, and event-driven architectures with Apache Kafka, Redis, Elasticsearch, and Kubernetes. I bring a strong background in MLOps and LLMOps, CI/CD automation, model deployment, observability, and multi-cloud platforms (AWS, Azure, GCP). I thrive on cross-functional collaboration to improve system scalability, reduce inference latency, and deliver reliable production AI solutions.

See more

Experience Level

Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
See more

Language

English
Fluent

Work Experience

AI Software Engineer at Carbyne
December 1, 2024 - Present
Engineered real-time AI microservices processing emergency audio streams with sub-400ms translation latency across distributed systems. Developed Retrieval-Augmented Generation (RAG) pipelines integrating LangChain, OpenAI APIs, and vector embeddings, improving contextual emergency-response summarization accuracy. Optimized streaming ASR workflows using Whisper-based models, Triton Inference Server, and Kubernetes, reducing inference latency during peak loads. Built AI agent orchestration services leveraging LangGraph, Redis, and Apache Kafka for scalable multi-step incident classification and routing. Implemented semantic search and vector retrieval pipelines using Pinecone and Elasticsearch to speed dispatcher workflows. Automated CI/CD deployment with Docker, Terraform, GitHub Actions, and Amazon EKS. Enhanced AI observability and model monitoring with Prometheus, Grafana, MLflow, and LangSmith dashboards.
AI Software Engineer at KPI T
January 1, 2022 - July 1, 2023
Created computer vision pipelines using PyTorch, OpenCV, and ONNX; optimized embedded AI inference workflows with NVIDIA TensorRT, CUDA, and quantization techniques to reduce ECU processing latency below 30 ms. Automated distributed PySpark data pipelines on Azure Databricks for multi-terabyte vehicle telemetry datasets, accelerating edge-case identification. Established RAG workflows integrating LangChain, vector embeddings, and historical defect logs to speed defect resolution. Integrated AI perception modules into ROS2 and AUTOSAR-based environments, improving interoperability across software-defined vehicle platforms. Automated model training and deployment pipelines using Docker, Jenkins, MLflow, and GitHub Actions, improving validation workflow efficiency. Supported functional safety and ISO 26262 automotive compliance requirements.
Software Engineer at Freshworks
January 1, 2021 - December 1, 2021
Architected scalable RESTful APIs using Java, Spring Boot, and MySQL, improving ticket-processing throughput across multi-tenant SaaS environments. Optimized Kafka-driven event processing pipelines and Redis caching layers, reducing customer ticket synchronization latency during peak traffic. Engineered Elasticsearch indexing workflows for customer-support platforms, accelerating ticket search response times across distributed microservices. Automated CI/CD deployment workflows with Docker, Jenkins, Kubernetes, and Terraform, improving deployment reliability and reducing production-release rollback incidents. Built asynchronous microservices integrating Kafka, Redis, and AWS services to enable reliable background processing for customer onboarding and workflow automation. Partnered with product managers, SREs, and frontend engineers to enhance API scalability, observability, and service reliability for global enterprise customers.

Education

Master’s in Computer Engineering at Florida Institute of Technology
January 11, 2030 - May 1, 2025
Bachelor’s in BSc-MECS at SR&BGNR COLLEGE
January 11, 2030 - June 1, 2022

Qualifications

Add your qualifications or awards here.

Industry Experience

Software & Internet, Professional Services, Telecommunications

Experience Level

Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Expert
See more