Senior Site Reliability Engineer - APAC

Tyk

IndiaremotePosted 20 days ago
Tyk logo

Skill Required

Site-Reliability-EngineeringDevOpsCloud-EngineerInfrastructure-EngineeringPlatform-EngineeringSenior-Site-Reliability-EngineerSite-Reliability-EngineerSite Reliability EngineerPenetration TestingNginxEngineeringKubernetesPrometheusnetworkingautomationObservabilityTerraformdesigningsimilar)buildingMongoDBetc.)TestNGTCP/IPRedisAzureLinuxCloudRustHelmAWSGCPDNSFulltime

Key highlights

  • Default remote with total flexibility in hours
  • Unlimited paid holidays for everyone
  • Employee share scheme
  • Generous maternity and paternity leave
  • Participating in on-call rotation: 16:00pm – 4:00am UTC
  • Building and maintaining global Tyk Cloud platform with multi-region and multi-cloud reach

Role overview

Tyk is an API management platform founded in 2015 that helps organizations connect their systems and services across various industries including retail, finance, telecoms, healthcare, and media. With offices in London (UK and Ontario), Atlanta, and Singapore, the company serves thousands of B2B users globally including brands like Lotte, Bell, T-Mobile, RBS, Capital One, and Vinci. Tyk operates as a total flexibility, default remote company with radical responsibility, offering unlimited paid holidays, and is on a mission to connect every system in the world through their platform.

Responsibilities

  • Maintaining global Tyk Cloud within SL(A/I/O)s you will help to define
  • Identifying reliability issues and working together with your squad to solve them
  • Identifying and introducing new metrics and building relevant dashboards
  • Participating in the on-call rotation
  • Working with your squad to expand multi-region and multi-cloud reach of the platform
  • Documenting operational knowledge
  • Conducting post-incident analysis
  • Automating common tasks
  • Be a key shaper and contributor to our continuous improvement agenda – be it the clarity of our user stories, how we estimate, communicate with other teams or customers – we expect this role to be advocate of continuous improvement
  • Reliability of our new global Tyk Cloud platform
  • Automation of operations and support
  • Writing and maintaining documentation on SRE processes and policies
  • Recommending and implementing ways of driving operational efficiency and driving down our cost to run, without impacting service
  • Assisting in penetration testing for Cloud through liaising with our provider, providing technical details, and environment setup
  • Incident management

Requirements

  • Strong collaboration skills
  • Launching and operating production scale kubernetes clusters
  • Designing and operating infrastructure on AWS and other providers
  • Operating MongoDB (or other document database) clusters
  • Operating Redis (or other key-value storage) clusters
  • Administering Linux servers
  • Maintaining distributed software
  • Operating Prometheus and Grafana
  • Operating logging collection and analysis systems
  • Participating in the on-call rotation(16:00pm – 4:00am UTC)
  • Kubernetes & containers (advanced)
  • AWS / EKS (advanced)
  • Linux (advanced)
  • Terraform and IaC in general (proficient)
  • Helm (proficient)
  • Go
  • MongoDB (or similar)
  • Redis (or similar)
  • Monitoring – prometheus, grafana, thanos (familiar)
  • Grasp of networking concepts (subnets, routing, peering, load balancing, NAT, etc.)
  • Common networking protocols (DNS, TCP/IP, HTTP, TLS, UDP)
  • Proactive, energetic, innovative and change oriented

Nice to have

  • GCP or Azure
  • Bare metal infrastructure engineering
  • API management experience
  • Large scale distributed storage management
  • Familiarity with Rancher
  • CKA/CKAD/CKS
  • Creating and delivering production software in Go language

Benefits

  • Everyone has unlimited paid holidays.
  • We have total flexibility in hours, as we believe creativity flows better when our people are given freedom to decide when they are most productive.
  • Employee share scheme
  • Generous maternity and paternity leave
  • Volunteering Days
  • Employee Wellbeing platform

Additional details

  • Tyk is an equal opportunities employer and we are determined to ensure that no applicant or employee receives less favourable treatment on the grounds of gender, age, disability, religion, belief, sexual orientation, marital status, or race, or is disadvantaged by conditions or requirements which cannot be shown to be justifiable.
  • Tyk was founded on the principle of offering flexibility and autonomy to our employees
  • Why we do this? Tyk was founded on the principle of offering flexibility and autonomy to our employees, we believe this allows our employees to achieve their best results. It also means we can build the best possible team, location and working hours are no barrier.
  • It's ok to screw up! We've found that it's often the 'stupid' or unexpected ideas that turn out to be the successful ones - so try it, at least we can say we have!
  • The only stupid idea, is the untested one!
  • Trust starts with you - make it count!
  • Assume best intent! We have each other's back - we're all on the same team. Think before you speak or act.
  • Make things better! Always try to leave things better than when you found them - change is constant, inevitable and embraced! Be that change we want to see.
  • Founded in 2015 with offices in London – UK, London – Ontario, Atlanta and Singapore
  • Brands using Tyk range from Lotte, Bell, T Mobile, to RBS, Capital One and Vinci
  • We have a varied user base hailing from every continent – even Antarctica.
Apply now