I’m an AI Engineer focused on building production-grade LLM applications, agentic AI, and multi-agent systems. I work across the full stack of AI engineering—designing agent workflows, implementing RAG and semantic search, and deploying scalable backends with FastAPI and Python using the OpenAI API and frameworks like LangChain and CrewAI. I enjoy optimizing real-world pipelines for reliability and performance, including building fault-tolerant microservices and long-term vector memory solutions. My goal is to turn complex AI requirements into modular, maintainable enterprise features that reduce manual effort and inference latency while supporting strong model evaluation and deployment cycles.

SonalY Jain

I’m an AI Engineer focused on building production-grade LLM applications, agentic AI, and multi-agent systems. I work across the full stack of AI engineering—designing agent workflows, implementing RAG and semantic search, and deploying scalable backends with FastAPI and Python using the OpenAI API and frameworks like LangChain and CrewAI. I enjoy optimizing real-world pipelines for reliability and performance, including building fault-tolerant microservices and long-term vector memory solutions. My goal is to turn complex AI requirements into modular, maintainable enterprise features that reduce manual effort and inference latency while supporting strong model evaluation and deployment cycles.

Available to hire

I’m an AI Engineer focused on building production-grade LLM applications, agentic AI, and multi-agent systems. I work across the full stack of AI engineering—designing agent workflows, implementing RAG and semantic search, and deploying scalable backends with FastAPI and Python using the OpenAI API and frameworks like LangChain and CrewAI.

I enjoy optimizing real-world pipelines for reliability and performance, including building fault-tolerant microservices and long-term vector memory solutions. My goal is to turn complex AI requirements into modular, maintainable enterprise features that reduce manual effort and inference latency while supporting strong model evaluation and deployment cycles.

See more

Experience Level

Expert
Expert
Expert
Expert
Intermediate
Intermediate
Intermediate
See more

Language

Work Experience

Associate AI Engineer at Teamarcs Technologies Pvt. Ltd.
September 1, 2025 - Present
Design and deploy production-grade LLM applications and AI automation tools using Python, LangChain, Hugging Face, FastAPI, and multi-agent frameworks (CrewAI). Optimize end-to-end inference pipelines and RAG architectures to reduce processing latency (~30%) and eliminate recurring production errors across deployed models. Build scalable AI microservices and integrate long-term vector memory using ChromaDB with fault-tolerant performance across distributed systems. Collaborate with cross-functional teams to translate business requirements into production AI features, supporting model evaluation and deployment cycles.
Data Science Intern at Zidio Development
June 1, 2024 - August 31, 2024
Built machine learning classification and prediction models using scikit-learn and TensorFlow, implementing full preprocessing and feature engineering pipelines across 3+ datasets. Integrated trained models into web applications to enable real-time inference for end users.
Data Analyst Intern at Cognorise Infotech
February 1, 2024 - April 30, 2024
Conducted exploratory data analysis and built business dashboards using Python, Pandas, and Matplotlib. Wrote optimized SQL queries for large-scale data extraction, transformation, and reporting, reducing manual reporting time significantly.
Backend Developer Intern at EiSystems — Technex, IIT BHU
June 1, 2023 - August 31, 2023
Developed RESTful APIs for AI/IoT projects using Python and Flask, implementing full CRUD operations with database integration. Deployed backend services to cloud platforms and ensured secure, authenticated API communication across distributed systems.

Education

B.Tech — Computer Science & Engineering (84%) at Teerthanker Mahaveer University, Moradabad
January 1, 2021 - January 1, 2025

Qualifications

LangChain for LLM Application Development
January 1, 2025 - September 13, 2026
AutoGen for Multi-Agent Systems
January 1, 2025 - September 13, 2026
Generative AI with CrewAI
January 1, 2025 - September 13, 2026
Data Structures & Algorithms with Python
January 1, 2024 - September 13, 2026
DBMS
January 1, 2023 - September 13, 2026

Industry Experience

Software & Internet, Computers & Electronics, Education, Professional Services
    AI-Based Online Examination & Proctoring System

    Developed a full-stack AI-powered proctoring platform ensuring academic integrity for remote examinations.

    ▸ Implemented real-time face recognition, gaze tracking, and motion detection using OpenCV and deep learning models to monitor student behavior during live exams
    ▸ Reduced cheating incidents by 30% through automated anomaly detection and instant alert generation
    ▸ Built React frontend with live video monitoring dashboard and Flask backend with RESTful APIs
    ▸ Integrated AWS S3 for secure media storage and Firebase for real-time event notifications
    ▸ Deployed on AWS with auto-scaling capabilities, improving system reliability and supporting concurrent users

    AI-Powered Resume Scanner & Scoring System

    Built an intelligent resume screening engine that automates candidate evaluation using semantic AI.

    ▸ Engineered a multi-stage parsing pipeline using GPT-based LLMs, spaCy NLP, and FAISS vector search to extract and semantically match skills, experience, and qualifications from PDF/DOCX resumes
    ▸ Implemented relevance scoring algorithm that ranks candidates against job descriptions with high accuracy
    ▸ Reduced manual recruiter screening time by 70% through automated ranking and skill-gap identification
    ▸ Designed a clean REST API interface via FastAPI for easy integration into existing HR workflows

    AI Survey Chatbot System

    Developed a context-aware conversational AI system for conducting intelligent, dynamic surveys.

    ▸ Built using OpenAI API and LangChain to deliver smooth, multi-turn survey conversations that adapt based on user responses
    ▸ Achieved 65% increase in user engagement compared to traditional static survey forms
    ▸ Designed conversation flows that handle edge cases, ambiguous answers, and topic branching intelligently
    ▸ Integrated response storage and analytics pipeline for post-survey insight generation

    AI Personal Assistant — Multi-Agent Autonomous System

    Built a fully autonomous multi-agent AI system capable of end-to-end task execution without human intervention.

    ▸ Designed a scalable multi-agent architecture using FastAPI, CrewAI, and OpenAI GPT-4, enabling agents to plan, reason, and collaborate autonomously
    ▸ Integrated long-term memory via ChromaDB, allowing the assistant to retain and recall context across sessions
    ▸ Built modular tool integrations: web search, PDF/document parsing, data analysis, email automation, and file generation
    ▸ Implemented microservice-based design for low-latency inference, fault tolerance, and independent agent scaling
    ▸ Supports fully context-aware responses for complex, multi-step real-world tasks.