We are seeking an Azure DevOps Engineer / Site Reliability Engineer (SRE) with 4–8 years of experience to manage, maintain, and optimize our CI/CD pipelines and Azure cloud infrastructure. The role involves collaborating with development teams to improve application stability, implementing Infrastructure as Code, and ensuring reliable production operations.
Responsibilities
- Manage, maintain, and optimize CI/CD pipelines.
- Deploy and support applications on Microsoft Azure.
- Monitor and maintain Azure cloud infrastructure.
- Define and track SLAs, SLOs, and Error Budgets.
- Automate operational tasks and reduce manual toil.
- Monitor application performance, availability, and reliability.
- Collaborate with development teams to improve application stability and release processes.
- Troubleshoot production issues and perform root cause analysis.
- Implement Infrastructure as Code (Terraform/Bicep/ARM preferred).
Requirements
- 4–8 years of relevant experience.
- Azure Cloud.
- Azure DevOps.
- CI/CD Pipelines.
- Kubernetes (AKS) / Docker.
- Terraform or ARM/Bicep.
- Git.
- PowerShell or Bash scripting.
- Monitoring tools (Azure Monitor, Prometheus, Grafana, Application Insights).
- SRE concepts (SLA, SLO, Error Budgets, Toil Reduction).
- Incident Management & Production Support.