Site Reliability Engineer (SRE)

Yopeso · via Himalayas ·

TypeFull-time job
LocationWorldwide
Posted1 hour ago
Yopeso has been developing a diverse range of software products, from large-scale applications to smaller solutions, for 19 years. With a growing team of over 250 employees across five locations, we are dedicated to fostering a culture of growth, transparency, and professionalism.
At Yopeso, we value authenticity, curiosity, and ambition. These values drive us to build strong connections within our community and with our partners, ensuring trust, integrity, and transparency in all our business practices. We strive to maintain the highest professional standards and continuously challenge ourselves to develop high-quality, high-performance, and secure software solutions.
Our approach is rooted in efficient collaboration among passionate professionals working in agile teams. Guided by curiosity and ambition, we strive to create meaningful, impactful products while staying true to our authentic selves.

What we offer:

Competitive remuneration

Remote work

24 days off per year and floating days

Private clinic health services, Regina Maria Medical Insurance

Flexible benefits through Up multi-benefits platform

Referral bonus scheme

Team events, online or at the office

Training and development opportunities with allocated budget

Professional Certifications

Knowledge-sharing context

Requirements
About the Project
We are looking for a Reliability / DevOps Engineer to join theteam, supporting a financial database platform built around business-critical database systems and related operational processes.
The role focuses on maintaining the stability, availability, performance, and reliability of large-scale databases, including environments containing terabytes of data.
You will work with database infrastructure, monitoring, upgrades, migrations, automation, and incident prevention in a highly reliability-focused environment.

Key Responsibilities

Monitor and maintain the reliability and availability of business-critical database systems.

Support databases containing terabytes of data and ensure stable operation in production environments.

Perform and support database version upgrades, patching, and maintenance activities.

Plan and execute data migrations with a strong focus on availability, integrity, and risk mitigation.

Investigate production issues, performance degradation, and reliability incidents.

Implement and improve monitoring, alerting, and observability for databases and supporting infrastructure.

Automate operational and infrastructure processes where appropriate.

Work with infrastructure and database teams to improve resilience, scalability, and operational efficiency.

Participate in root cause analysis and contribute to preventive measures following incidents.

Maintain and improve operational procedures, runbooks, and reliability practices.

Required Skills

Hands-on experience in DevOps, Site Reliability Engineering, Infrastructure Engineering, or Database Operations.

Good understanding of database management and database reliability concepts.

Experience supporting large or business-critical production systems.

Practical experience with monitoring and observability tools.

Understanding of performance monitoring, alerting, troubleshooting, and incident management.

Experience working with data migrations and/or database upgrades.

Good knowledge of Linux and infrastructure troubleshooting.

Ability to work independently and take ownership of production reliability topics.

Nice to Have

Experience with AWS.

Experience with Terraform or other Infrastructure-as-Code tools.

Experience with enterprise monitoring and observability platforms.

Previous experience working with Oracle Databases.

Experience managing databases at terabyte scale.

Background in Database Reliability Engineering (DBRE) or SRE environments.

Experience supporting systems in the financial services or other highly regulated industries.

Originally posted on Himalayas
site-reliability-engineer devops-engineer sre database-reliability-engineer infrastructure-engineer site-reliability-engineering-(sre) site-reliability-operations-engineer devops-site-reliability-engineer site-reliability-engineering-jobs senior-site-reliability-engineer staff-site-reliability-engineer-(sre) site-reliability-engineering site-reliability-engineer-ii
Apply on Himalayas →

Job sourced from Himalayas. Applications happen directly on the original platform — we never collect your data.