Posted today · be early
Senior Site Reliability Engineer
N-iX
WorldwideremotePosted today
Skill Required
Site-Reliability-EngineeringSRE-EngineerCloud-EngineerDevOps-EngineerPlatform-EngineeringSenior-Site-Reliability-EngineerSenior-Site-Reliability-Engineering-ArchitectPrincipal-Site-Reliability-EngineerSite-Reliability-Engineering-LeadDevOps-Site-Reliability-EngineerSite Reliability EngineerCI/CDSystem DesigntroubleshootingGCPIstioEngineeringR ProgrammingJavaScriptautomationdevelopingObservabilityData StructuresTerraformdesigningetc.)TestNGPythondesignDevOpsC++AzureLinuxCloudJavaRubyFulltime
Key highlights
- Role: Senior Site Reliability Engineer
- Experience: Extensive expertise in software development/testing, DevOps, and SRE
- Key Tech: Azure, Kubernetes, Terraform
- Benefit: Flexible working format (remote, office-based or flexible)
- Compensation: Competitive salary and good compensation package
Role overview
We are looking for a Senior Site Reliability Engineer to work for an innovative hospitality company utilizing cutting-edge technologies. The role involves working within an international team to develop brilliant products on reliable and resilient systems running in Azure and traditional data centers. Site Reliability Engineering (SRE) at this company combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SREs ensure that internally critical and externally-visible cloud services have reliability, uptime appropriate to customer needs, and a fast rate of improvement, while maintaining a watchful eye on capacity and performance.
Responsibilities
- Develop and improve the whole lifecycle of services
- Establish and improve monitoring capabilities to reduce outage frequency and duration
- Create sustainable systems through automation and uplifts
- Develop and scale systems sustainably through mechanisms such as automation, and evolve systems by pushing for changes that improve reliability and velocity.
- Lead designs of major software components, systems, and features to improve the availability, scalability, latency, and efficiency of our services
- Analyze and support services before they go live via system design consulting, developing software platforms and frameworks, capacity planning
- Conduct post-incident analysis and reviews with an attitude of continuous improvement
- Manage the complex challenges of scale that are unique to the project while using your expertise in coding, algorithms, complexity analysis, and large-scale system design
- Provide scalable, reliable, durable, and secure services using a customer-first approach while innovating technically
- Understand our customer needs and how we can meet them
Requirements
- Ideally, strong experience in Azure Services and capabilities, but other cloud services (AWS, Google Cloud Platform etc.) will be considered
- Confidence and strong experience with Kubernetes
- Recent and fluent Terraform
- Extensive expertise in software development/testing, development operations, and site reliability engineering
- Experience of Unix/Linux administration
- Experience with Continuous Integration and Deployment (CI/CD) and release orchestration and Configuration Management of VMs
- Cloud-agnostic approach, with flexibility to work across various cloud platforms
- Experience programming in one or more of the following languages: C#, C++, Java, Python, JavaScript, Go, Perl, or Ruby
Nice to have
- Chef platform experience
- An appreciation of systems internals (e.g., filesystems, system calls)
- Bachelor's degree in Computer Science, similar technical field of study, or equivalent practical experience
- Experience in distributed systems, storage systems, or databases
- Experience designing, analyzing, and troubleshooting large-scale distributed systems
- Systematic problem-solving approach, combined with excellent communication skills and a sense of ownership and drive
- Experience in configuring application monitoring with Azure Monitor and Application Insight
- Experience with Service Mesh
- Previous experience as a DevOps engineer is preferred
Benefits
- Flexible working format - remote, office-based or flexible
- A competitive salary and good compensation package
- Personalized career growth
- Professional development tools (mentorship program, tech talks and trainings, centers of excellence, and more)
- Active tech communities with regular knowledge sharing
- Education reimbursement
- Memorable anniversary presents
- Corporate events and team buildings
- Other location-specific benefits
Additional details
- Benefits are not applicable for freelancers
- Originally posted on Himalayas