Site Reliability Engineer – SRE

Posted 1ds ago

Employment Information

Education
Salary
Experience
Job Type

Report this job

Job expired or something wrong with this job?

Job Description

Site Reliability Engineer maintaining Yopeso’s business-critical financial database platform. Improving database reliability, observability, migrations, upgrades, and production resilience.

Responsibilities:

  • Monitor and maintain the reliability and availability of business-critical database systems
  • Support databases containing terabytes of data and ensure stable operation in production environments
  • Perform and support database version upgrades, patching, and maintenance activities
  • Plan and execute data migrations with a strong focus on availability, integrity, and risk mitigation
  • Investigate production issues, performance degradation, and reliability incidents
  • Implement and improve monitoring, alerting, and observability for databases and supporting infrastructure
  • Automate operational and infrastructure processes where appropriate
  • Work with infrastructure and database teams to improve resilience, scalability, and operational efficiency
  • Participate in root cause analysis and contribute to preventive measures following incidents
  • Maintain and improve operational procedures, runbooks, and reliability practices

Requirements:

  • Hands-on experience in DevOps, Site Reliability Engineering, Infrastructure Engineering, or Database Operations
  • Good understanding of database management and database reliability concepts
  • Experience supporting large or business-critical production systems
  • Practical experience with monitoring and observability tools
  • Understanding of performance monitoring, alerting, troubleshooting, and incident management
  • Experience working with data migrations and/or database upgrades
  • Good knowledge of Linux and infrastructure troubleshooting
  • Ability to work independently and take ownership of production reliability topics
  • Experience with AWS
  • Experience with Terraform or other Infrastructure-as-Code tools
  • Experience with enterprise monitoring and observability platforms
  • Previous experience working with Oracle Databases
  • Experience managing databases at terabyte scale
  • Background in Database Reliability Engineering (DBRE) or SRE environments
  • Experience supporting systems in the financial services or other highly regulated industries

Benefits:

  • Competitive remuneration
  • Remote work
  • 24 days off per year and floating days
  • Private clinic health services, Regina Maria Medical Insurance
  • Flexible benefits through Up multi-benefits platform
  • Referral bonus scheme
  • Team events, online or at the office
  • Training and development opportunities with allocated budget
  • Professional Certifications
  • Knowledge-sharing context