I’m Siddhanth Balakrishnan, an AI engineer based in Chennai who loves building end-to-end pipelines that turn raw data into usable, high-quality outputs. At OO Studio, I’ve worked across the full workflow—from designing dataset pipelines for video training, to training character LoRAs for repeatable identity control, to evaluating and chaining modern video generation models for longer clips. I also enjoy making these systems practical and reliable in real-time settings. I’ve built production-grade components like lip sync, restoration, upscaling/colorization workflows in ComfyUI, and a real-time multi-speaker dubbing pipeline for live streams (with forced alignment improvements). I ship with strong deployment discipline, using Docker, FastAPI, gRPC, and monitoring with Prometheus/Grafana, while continuously pushing performance and quality in AI video and speech/audio systems.

Siddhanth Balakrishnan

I’m Siddhanth Balakrishnan, an AI engineer based in Chennai who loves building end-to-end pipelines that turn raw data into usable, high-quality outputs. At OO Studio, I’ve worked across the full workflow—from designing dataset pipelines for video training, to training character LoRAs for repeatable identity control, to evaluating and chaining modern video generation models for longer clips. I also enjoy making these systems practical and reliable in real-time settings. I’ve built production-grade components like lip sync, restoration, upscaling/colorization workflows in ComfyUI, and a real-time multi-speaker dubbing pipeline for live streams (with forced alignment improvements). I ship with strong deployment discipline, using Docker, FastAPI, gRPC, and monitoring with Prometheus/Grafana, while continuously pushing performance and quality in AI video and speech/audio systems.

Available to hire

I’m Siddhanth Balakrishnan, an AI engineer based in Chennai who loves building end-to-end pipelines that turn raw data into usable, high-quality outputs. At OO Studio, I’ve worked across the full workflow—from designing dataset pipelines for video training, to training character LoRAs for repeatable identity control, to evaluating and chaining modern video generation models for longer clips.

I also enjoy making these systems practical and reliable in real-time settings. I’ve built production-grade components like lip sync, restoration, upscaling/colorization workflows in ComfyUI, and a real-time multi-speaker dubbing pipeline for live streams (with forced alignment improvements). I ship with strong deployment discipline, using Docker, FastAPI, gRPC, and monitoring with Prometheus/Grafana, while continuously pushing performance and quality in AI video and speech/audio systems.

See more

Language

Tamil
Fluent
English
Advanced
Hindi
Advanced
French
Advanced
German
Intermediate

Work Experience

AI Engineer at OO Studio
August 1, 2025 - Present
Built an LTX training dataset pipeline that splits raw video into shots, isolates clean character shots, transcribes speech with an ASR model, and generates captions in the exact format (resolution and frame rate) required for training. Trained a character LoRA on RunPod to reproduce a person’s face and voice from a small sample, enabling direct identity control and a repeatable path for training LoRAs for new subjects. Developed lip sync, colorisation, upscaling, and restoration workflows in ComfyUI and deployed them on RunPod, comparing multiple VSR/colorization models and selecting SeedVR2. Evaluated MiniMax H3 and implemented continuous generation chaining to extend clips to ~45 seconds. Shipped ScriptFlow (FastAPI + React) that turns a screenplay into a storyboard using an LLM stage plus an image generation pipeline with live WebSocket progress. Built real-time multi-speaker dubbing for live streams with per-speaker ASR→translation→TTS lanes, then time-aligned/fitted transla
AI Engineer (Internship) at ANTZ.AI
October 1, 2023 - December 1, 2023
Built an AI interviewer using FAISS-based RAG to ask job-specific questions extracted from job descriptions and handle candidate follow-ups. Added a feedback module that scored candidates and returned strengths, weaknesses, and sample answers. Developed a resume-screening tool that filtered applications against job-description requirements.

Education

B.Tech, Computer Science Engineering (AI & ML) at Sri Ramachandra Institute of Higher Education and Research
January 1, 2021 - January 1, 2025
Higher Secondary Education (TN State Board) at MCTM Chidambaram Chettyar
January 1, 2019 - January 1, 2021

Qualifications

Add your qualifications or awards here.

Industry Experience

Software & Internet, Media & Entertainment, Computers & Electronics, Education