I’m a Senior Site Reliability Engineer and platform engineer with 11 years of experience designing, automating, and operating large-scale cloud infrastructure on AWS and Azure. I focus on making operations faster and safer—reducing investigation time, improving reliability, and lowering cloud costs through repeatable infrastructure, strong change governance, and observability that leads to action.
Recently, I led an Agentic AI for Ops initiative delivering an SRE Agent POC using LLM agents and MCP integrations to triage incidents end-to-end across PagerDuty, Grafana/Prometheus metrics, Kubernetes health signals, and MongoDB/Elasticsearch logs. I also drive incident response and reliability improvements by owning the full lifecycle (detect, mitigate, communicate, and follow-through), mentoring engineers, and building secure CI/CD and cloud-native platforms that scale across multi-region environments.
Skills
Experience Level
Work Experience
Education
Qualifications
Industry Experience
Skills
Experience Level
Hire a DevOps Developer
We have the best devops developer experts on Twine. Hire a devops developer in Hyderabad today.