Posted today · be early
Reliability Engineer / Devops - Database Platform
Yopeso
WorldwideremotePosted today
Skill Required
Site-Reliability-EngineerDevOps-EngineerDatabase-AdministratorDatabase-EngineerInfrastructure-EngineerSenior-Database-Reliability-EngineerDevOps EngineerPlatform EngineerDevOpstroubleshootingObservabilityEngineeringautomationdevelopingTerraformOracleLinuxAgileRustAWSandFulltime
Key highlights
- Financial database platform role supporting business-critical systems
- Terabyte-scale database management required
- 32 calendar days of vacation provided
- Professional development plan with mentorship included
- Remote work option available
- Experience with large or business-critical production systems required
Role overview
Yopeso, a software development company with over 20 years of experience and a team of 300+ professionals across five locations, is seeking a Reliability / DevOps Engineer to join our team. The role involves supporting a financial database platform built around business-critical database systems, focusing on maintaining the stability, availability, performance, and reliability of large-scale databases, including environments with terabytes of data. Our approach emphasizes efficient collaboration in agile teams, driven by curiosity and ambition to create meaningful, high-quality software solutions.
Responsibilities
- Monitor and maintain the reliability and availability of business-critical database systems.
- Support databases containing terabytes of data and ensure stable operation in production environments.
- Perform and support database version upgrades, patching, and maintenance activities.
- Plan and execute data migrations with a strong focus on availability, integrity, and risk mitigation.
- Investigate production issues, performance degradation, and reliability incidents.
- Implement and improve monitoring, alerting, and observability for databases and supporting infrastructure.
- Automate operational and infrastructure processes where appropriate.
- Work with infrastructure and database teams to improve resilience, scalability, and operational efficiency.
- Participate in root cause analysis and contribute to preventive measures following incidents.
- Maintain and improve operational procedures, runbooks, and reliability practices.
Requirements
- Hands-on experience in DevOps, Site Reliability Engineering, Infrastructure Engineering, or Database Operations.
- Good understanding of database management and database reliability concepts.
- Experience supporting large or business-critical production systems.
- Practical experience with monitoring and observability tools.
- Understanding of performance monitoring, alerting, troubleshooting, and incident management.
- Experience working with data migrations and/or database upgrades.
- Good knowledge of Linux and infrastructure troubleshooting.
- Ability to work independently and take ownership of production reliability topics.
Nice to have
- Experience with AWS.
- Experience with Terraform or other Infrastructure-as-Code tools.
- Experience with enterprise monitoring and observability platforms.
- Previous experience working with Oracle Databases.
- Experience managing databases at terabyte scale.
- Background in Database Reliability Engineering (DBRE) or SRE environments.
- Experience supporting systems in the financial services or other highly regulated industries.
Benefits
- Competitive remuneration
- Remote work
- Sports/leisure benefit
- 20 sick leave days paid at 100%
- 32 calendar days of vacation
- Team events, online, at the office, or outside
- Professional development plan with guidance and mentorship
- Training and development opportunities with an allocated budget
- Professional Certifications
- Optional medical insurance