I am a Site Reliability Engineer with hands-on experience designing, automating, and maintaining production systems across AWS and Azure. I focus on reliability through observability, automation, and operational excellence, leveraging Kubernetes, IaC, and CI/CD pipelines to deliver scalable, zero-downtime SaaS releases. I thrive in collaborative, Agile environments and enjoy turning complex infrastructure problems into repeatable, robust solutions. I enjoy partnering with development teams to bake in production readiness and post-incident learning, reduce toil, and improve SLO adherence. I write Python and Bash automation to simplify ops tasks and visibility, and I am passionate about building resilient cloud platforms and driving incident reduction through proactive monitoring and strong ITIL-aligned processes.

Darshan Gowda A.R.

I am a Site Reliability Engineer with hands-on experience designing, automating, and maintaining production systems across AWS and Azure. I focus on reliability through observability, automation, and operational excellence, leveraging Kubernetes, IaC, and CI/CD pipelines to deliver scalable, zero-downtime SaaS releases. I thrive in collaborative, Agile environments and enjoy turning complex infrastructure problems into repeatable, robust solutions. I enjoy partnering with development teams to bake in production readiness and post-incident learning, reduce toil, and improve SLO adherence. I write Python and Bash automation to simplify ops tasks and visibility, and I am passionate about building resilient cloud platforms and driving incident reduction through proactive monitoring and strong ITIL-aligned processes.

Available to hire

I am a Site Reliability Engineer with hands-on experience designing, automating, and maintaining production systems across AWS and Azure. I focus on reliability through observability, automation, and operational excellence, leveraging Kubernetes, IaC, and CI/CD pipelines to deliver scalable, zero-downtime SaaS releases. I thrive in collaborative, Agile environments and enjoy turning complex infrastructure problems into repeatable, robust solutions.

I enjoy partnering with development teams to bake in production readiness and post-incident learning, reduce toil, and improve SLO adherence. I write Python and Bash automation to simplify ops tasks and visibility, and I am passionate about building resilient cloud platforms and driving incident reduction through proactive monitoring and strong ITIL-aligned processes.

See more

Experience Level

Expert
Expert
Expert
Expert
Expert
Expert
Expert

Language

English
Fluent

Work Experience

DevOps / SRE Engineer at TestYantra Software Solution
March 1, 2022 - October 1, 2023
Led infrastructure as code efforts using CloudFormation and Terraform to automate provisioning of VPCs, EC2s, RDS, S3, and more, increasing deployment speed by 60%. Managed containerized applications on Azure Kubernetes Service, troubleshooting pod failures, network policies, and optimizing resources for production workloads. Maintained production-grade AWS environments with 99.99% SLA adherence, enabling zero-downtime SaaS releases through proactive monitoring and instrumentation. Developed Python and Bash automation scripts for infrastructure management and data validation, reducing manual toil by 45%. Built and maintained CI/CD pipelines using GitHub Actions and Jenkins for automated testing and deployment, improving deployment reliability by 35%. Owned ITIL-aligned Problem Management, conducting post-incident root cause analyses and implementing preventive measures. Configured CloudWatch alarms and dashboards for EC2 CPU, memory, disk usage, and application logs, enabling proactive
DevOps Engineer - Intern at SkillRary (c/o TestYantra Software Solution)
June 1, 2021 - February 1, 2022
Configured CloudWatch alarms and dashboards for EC2 metrics and application logs, enabling proactive remediation and 40% fewer incidents through alert tuning. Implemented proactive monitoring using New Relic with SLOs and automated health checks to detect service degradation before customer impact, reducing mean time to detect by 35%. Worked with Ansible scripts for troubleshooting, automation, and patch configuration in cloud environments, and documented infrastructure workflows to support onboarding and knowledge sharing. Managed AWS Cloud production environments ensuring uptime with proactive monitoring and automated health checks that reduced MTTR by 35%. Monitored build and deployment health via Jira, and developed dashboards with Prometheus and Grafana for proactive observability. Automated backup and disaster recovery workflows with AWS Backup, S3 lifecycle policies, and Glacier vaults, significantly reducing RTO.

Education

Master in DevOps (MSc) at ATU, Letterkenny, Ireland
January 1, 2024 - January 1, 2025
Bachelor Of Computer Application (BCA) at University of Mysore, India
January 1, 2018 - January 1, 2021
Master in DevOps (MSc) at ATU, Letterkenny, Ireland
January 1, 2024 - January 1, 2025
Bachelor of Computer Applications (BCA) at University of Mysore, India
January 1, 2018 - January 1, 2021
Master in DevOps (MSc) at ATU, Letterkenny, Ireland
January 1, 2024 - January 1, 2025
Bachelor Of Computer Application (BCA) at University of Mysore, India
January 1, 2018 - January 1, 2021

Qualifications

AWS Certified Solutions Architect – Associate
January 11, 2030 - April 20, 2026
HashiCorp Certified: Terraform Associate (003)
January 11, 2030 - April 20, 2026
GitHub Foundations certification
January 11, 2030 - April 20, 2026
CKA - Certified Kubernetes Administrator
January 11, 2030 - April 20, 2026
AWS Certified Solutions Architect – Associate
January 11, 2030 - April 20, 2026
HashiCorp Certified: Terraform Associate (003)
January 11, 2030 - April 20, 2026
GitHub Foundations certification
January 11, 2030 - April 20, 2026
CKA - Certified Kubernetes Administrator
January 11, 2030 - April 20, 2026
AWS Certified Solutions Architect – Associate
January 11, 2030 - April 24, 2026
HashiCorp Certified: Terraform Associate
January 11, 2030 - April 24, 2026
GitHub Foundations certification
January 11, 2030 - April 24, 2026
CKA - Certified Kubernetes Administrator
January 11, 2030 - April 24, 2026

Industry Experience

Software & Internet, Professional Services, Other