Site Reliability Engineer – SRE
Posted 1ds ago
Employment Information
Report this job
Job expired or something wrong with this job?
Job Description
Site Reliability Engineer maintaining Yopeso’s business-critical financial database platform. Improving database reliability, observability, migrations, upgrades, and production resilience.
Responsibilities:
- Monitor and maintain the reliability and availability of business-critical database systems
- Support databases containing terabytes of data and ensure stable operation in production environments
- Perform and support database version upgrades, patching, and maintenance activities
- Plan and execute data migrations with a strong focus on availability, integrity, and risk mitigation
- Investigate production issues, performance degradation, and reliability incidents
- Implement and improve monitoring, alerting, and observability for databases and supporting infrastructure
- Automate operational and infrastructure processes where appropriate
- Work with infrastructure and database teams to improve resilience, scalability, and operational efficiency
- Participate in root cause analysis and contribute to preventive measures following incidents
- Maintain and improve operational procedures, runbooks, and reliability practices
Requirements:
- Hands-on experience in DevOps, Site Reliability Engineering, Infrastructure Engineering, or Database Operations
- Good understanding of database management and database reliability concepts
- Experience supporting large or business-critical production systems
- Practical experience with monitoring and observability tools
- Understanding of performance monitoring, alerting, troubleshooting, and incident management
- Experience working with data migrations and/or database upgrades
- Good knowledge of Linux and infrastructure troubleshooting
- Ability to work independently and take ownership of production reliability topics
- Experience with AWS
- Experience with Terraform or other Infrastructure-as-Code tools
- Experience with enterprise monitoring and observability platforms
- Previous experience working with Oracle Databases
- Experience managing databases at terabyte scale
- Background in Database Reliability Engineering (DBRE) or SRE environments
- Experience supporting systems in the financial services or other highly regulated industries
Benefits:
- Competitive remuneration
- Remote work
- 24 days off per year and floating days
- Private clinic health services, Regina Maria Medical Insurance
- Flexible benefits through Up multi-benefits platform
- Referral bonus scheme
- Team events, online or at the office
- Training and development opportunities with allocated budget
- Professional Certifications
- Knowledge-sharing context

















