Posted today · be early

Devops Engineer

Ziventra

WorldwideremotePosted today
Ziventra logo

Skill Required

DevOps-EngineerPlatform-EngineerSite-Reliability-EngineerCloud-EngineerDevOps-Software-EngineerCloud-DevOps-EngineerProdOps-EngineerDevOps-Automation-EngineerInfrastructure-DevOps-EngineerAWS-DevOps-EngineerTechOps-EngineerCloudOps-EngineerDevOps EngineerDevOpsSOCDockertroubleshootingMicroservicesObservabilityCD pipelinesEngineeringR ProgrammingKubernetesPrometheusautomationTerraformdesigningsecuritybuildingetc.)GolangArgoCDdesignCI/CDCloudDesign PatternsAWSIAMContract

Key highlights

  • Experience level: 4–8 years
  • Key requirement: Strong Golang, Kubernetes, and AWS experience
  • Key requirement: Expertise in Terraform (IaC)

Role overview

We are looking for a highly skilled DevOps / Platform Engineer with strong hands-on experience in Golang, Kubernetes, and AWS to design, build, and operate scalable cloud-native platforms. The ideal candidate will be responsible for CI/CD automation, infrastructure provisioning, container orchestration, monitoring, and incident management in production environments.

Responsibilities

  • Design, deploy, and manage highly available and scalable infrastructure on AWS
  • Implement Infrastructure as Code (IaC) using Terraform for repeatable and secure deployments
  • Manage and optimize Kubernetes clusters (EKS preferred), including upgrades, scaling, and security hardening
  • Design and implement end-to-end CI/CD pipelines for microservices and cloud-native applications
  • Automate build, test, security scans, and deployment workflows
  • Ensure best practices for pipeline reliability, performance, and security
  • Build, optimize, and manage containerized applications using Docker
  • Implement Kubernetes best practices including deployments, services, ingress, HPA, and resource management
  • Enforce container security and image lifecycle management
  • Implement and maintain monitoring and observability solutions (Prometheus, Grafana, CloudWatch, etc.)
  • Define and monitor SLIs, SLOs, and SLAs
  • Proactively identify performance bottlenecks and reliability risks
  • Participate in on-call rotations and handle production incidents
  • Lead incident response, troubleshooting, and recovery efforts
  • Perform Root Cause Analysis (RCA) and implement preventive measures
  • Work closely with development, security, and architecture teams
  • Promote DevOps, SRE, and cloud-native best practices across teams
  • Contribute to platform documentation and operational runbooks

Requirements

  • Strong programming experience in Golang
  • Strong hands-on experience with Kubernetes (production environments)
  • Strong hands-on experience with AWS (EKS, EC2, IAM, VPC, S3, Load Balancers, CloudWatch)
  • Proven experience designing and implementing CI/CD pipelines
  • Expertise in Infrastructure as Code, preferably Terraform
  • Deep understanding of containerization and orchestration best practices
  • Experience with monitoring, logging, and observability tools
  • Solid experience in incident management and root cause analysis
  • Bachelor’s degree in Computer Science, Engineering, or equivalent practical experience
  • 4–8 years of experience in DevOps / Platform / Cloud Engineering
  • Strong problem-solving and troubleshooting skills
  • Excellent communication and collaboration abilities

Nice to have

  • Experience building or supporting cloud-native microservices architectures
  • Knowledge of cost optimization strategies in AWS
  • Exposure to Site Reliability Engineering (SRE) principles
  • Familiarity with security scanning tools and DevSecOps practices
  • Experience with GitOps tools (ArgoCD, Flux)

Additional details

  • Originally posted on Himalayas
Apply now