Posted today · be early
Devops Engineer
Ziventra
WorldwideremotePosted today
Skill Required
DevOps-EngineerPlatform-EngineerSite-Reliability-EngineerCloud-EngineerDevOps-Software-EngineerCloud-DevOps-EngineerProdOps-EngineerDevOps-Automation-EngineerInfrastructure-DevOps-EngineerAWS-DevOps-EngineerTechOps-EngineerCloudOps-EngineerDevOps EngineerDevOpsSOCDockertroubleshootingMicroservicesObservabilityCD pipelinesEngineeringR ProgrammingKubernetesPrometheusautomationTerraformdesigningsecuritybuildingetc.)GolangArgoCDdesignCI/CDCloudDesign PatternsAWSIAMContract
Key highlights
- Experience level: 4–8 years
- Key requirement: Strong Golang, Kubernetes, and AWS experience
- Key requirement: Expertise in Terraform (IaC)
Role overview
We are looking for a highly skilled DevOps / Platform Engineer with strong hands-on experience in Golang, Kubernetes, and AWS to design, build, and operate scalable cloud-native platforms. The ideal candidate will be responsible for CI/CD automation, infrastructure provisioning, container orchestration, monitoring, and incident management in production environments.
Responsibilities
- Design, deploy, and manage highly available and scalable infrastructure on AWS
- Implement Infrastructure as Code (IaC) using Terraform for repeatable and secure deployments
- Manage and optimize Kubernetes clusters (EKS preferred), including upgrades, scaling, and security hardening
- Design and implement end-to-end CI/CD pipelines for microservices and cloud-native applications
- Automate build, test, security scans, and deployment workflows
- Ensure best practices for pipeline reliability, performance, and security
- Build, optimize, and manage containerized applications using Docker
- Implement Kubernetes best practices including deployments, services, ingress, HPA, and resource management
- Enforce container security and image lifecycle management
- Implement and maintain monitoring and observability solutions (Prometheus, Grafana, CloudWatch, etc.)
- Define and monitor SLIs, SLOs, and SLAs
- Proactively identify performance bottlenecks and reliability risks
- Participate in on-call rotations and handle production incidents
- Lead incident response, troubleshooting, and recovery efforts
- Perform Root Cause Analysis (RCA) and implement preventive measures
- Work closely with development, security, and architecture teams
- Promote DevOps, SRE, and cloud-native best practices across teams
- Contribute to platform documentation and operational runbooks
Requirements
- Strong programming experience in Golang
- Strong hands-on experience with Kubernetes (production environments)
- Strong hands-on experience with AWS (EKS, EC2, IAM, VPC, S3, Load Balancers, CloudWatch)
- Proven experience designing and implementing CI/CD pipelines
- Expertise in Infrastructure as Code, preferably Terraform
- Deep understanding of containerization and orchestration best practices
- Experience with monitoring, logging, and observability tools
- Solid experience in incident management and root cause analysis
- Bachelor’s degree in Computer Science, Engineering, or equivalent practical experience
- 4–8 years of experience in DevOps / Platform / Cloud Engineering
- Strong problem-solving and troubleshooting skills
- Excellent communication and collaboration abilities
Nice to have
- Experience building or supporting cloud-native microservices architectures
- Knowledge of cost optimization strategies in AWS
- Exposure to Site Reliability Engineering (SRE) principles
- Familiarity with security scanning tools and DevSecOps practices
- Experience with GitOps tools (ArgoCD, Flux)
Additional details
- Originally posted on Himalayas