DevOps Engineer
Petals Careers Private Limited
Role tags
Tech stack mentioned
Role overview
Formatting this description...
Senior DevOps Engineer Role Overview We are looking for a Senior DevOps Engineer with 5+ years of experience in building and operating production-grade infrastructure. This role focuses on automation, scalability, reliability, and developer productivity. You will design and manage cloud infrastructure, build efficient CI/CD pipelines, improve observability, and drive engineering best practices across the organization. Responsibilities Design, build, and manage scalable, secure, and highly available cloud infrastructure. Develop and maintain reliable CI/CD pipelines for fast and safe deployments. Automate infrastructure provisioning and operational workflows using Infrastructure as Code (IaC). Implement monitoring, logging, alerting, and incident management to improve system reliability. Define and drive SLOs, SLIs, and reliability best practices. Strengthen infrastructure security through secrets management, access controls, and compliance practices. Troubleshoot and resolve complex production issues with a focus on root cause analysis. Collaborate with engineering teams to improve deployment processes and developer productivity. Mentor engineers and promote DevOps, automation, and operational excellence. Requirements Experience 5+ years of experience in DevOps, Site Reliability Engineering (SRE), or Platform Engineering. Proven experience managing production infrastructure at scale. Cloud & Infrastructure Strong hands-on experience with AWS, GCP, or Azure. Experience with Docker, Kubernetes, or similar container orchestration platforms. CI/CD & Automation Experience building and maintaining CI/CD pipelines using GitHub Actions, GitLab CI, Jenkins, or similar tools. Strong scripting skills in Bash, Python, or a similar language. Infrastructure as Code Hands-on experience with Terraform, CloudFormation, or equivalent IaC tools. Strong understanding of Git-based infrastructure workflows. Observability & Reliability Experience with monitoring, logging, and alerting tools. Good understanding of distributed systems, scalability, and reliability engineering principles. Skills Strong problem-solving and debugging skills for production environments. Automation-first mindset with a strong sense of ownership. Excellent communication and collaboration skills. Show more