Available to hire
I am Kirill Kuklin, a Canada-based Senior Site Reliability Engineer with 12 years of international experience designing, implementing, and maintaining highly available infrastructure and distributed systems.
I specialize in containerization (Kubernetes, Docker), Infrastructure-as-Code (Terraform), robust monitoring, alerting, and leading incident response for mission-critical applications. I thrive in 24/7 on-call operations, post-incident analyses, and mentoring teams in SRE best practices.
Skills
Experience Level
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Language
English
Fluent
Work Experience
Senior Site Reliability Engineer (Security and Compliance) at Microsoft
June 1, 2021 - November 6, 2025Supported Microsoft Defender family software by applying Linux/macOS knowledge (auditd, eBPF, netext). Migrated critical tools from a Linux-based VM (Docker Compose) to AKS, resolving space limitations and frequent downtimes. Re-architected deployment using Docker images and persistent Azure Blob storage for scalable logging and data retention. Troubleshot Defender products on Linux and macOS (performance and availability) and documented procedures in playbooks. Developed Python-based ETL pipelines (GitHub Actions) to extract content and metadata from ADO articles, leveraging Azure OpenAI to summarize articles and generate visit graphs in Grafana. Architected scalable CI/CD pipelines, and used machine learning models to identify and defer low-quality or misrouted incidents, reducing triage load. Maintained 24/7 on-call responsibilities for Defender infrastructure, led incident response for P1/P2 outages, and coordinated cross-team efforts to restore service within SLA targets. Designed
Expert Customer Solutions Engineer at Veeam
June 1, 2021 - June 1, 2021Operated enterprise-scale backup infrastructure for Fortune 500 customers with multi-petabyte data volumes, maintaining 99.9% availability and strict RTO/RPO SLA compliance. Developed automation using Python, PowerShell, and Microsoft Graph API for M365 tenant management; built automated monitoring and incident correlation systems using Python pipelines and AI-based summarization; conducted performance assessments and capacity planning; mentored a team of engineers on backup architecture and incident handling best practices.
L2 Technical Support Engineer at Crossover
May 1, 2016 - May 1, 2016Operated multi-cloud infrastructure across AWS and Azure platforms, performing deep-dive troubleshooting through code analysis, database queries, and log correlation to resolve complex system issues. Built automation and integration solutions using PowerShell scripts and REST APIs for data export/import workflows, demonstrating early foundation in infrastructure-as-code practices. Diagnosed and resolved distributed system failures including serverless components (AWS Lambda), API services, and data pipeline availability issues. Implemented observability and monitoring solutions using the ELK stack for log aggregation, query optimization, and operational dashboards.
Education
Associate Certificate in Computer Systems Technology at British Columbia Institute of Technology
January 11, 2030 - November 6, 2025Qualifications
Industry Experience
Software & Internet, Professional Services
Skills
Experience Level
Expert
Expert
Expert
Expert
Expert
Expert
Expert
Hire a Full Stack Developer
We have the best full stack developer experts on Twine. Hire a full stack developer in Vancouver today.