Remote
Site Reliability Engineer
About this role
Who We Are Moniepoint Inc. is Africa’s all-in-one financial platform, helping 20 million businesses and individuals access seamless payments, banking, credit, cross-border, and business management tools each month. As Nigeria’s largest merchant acquirer, we power most of the country’s point-of-sale (POS) transactions. Through our subsidiaries, Moniepoint Inc. processes over $250 billion in digital payment transaction value annually.
What We Do At Moniepoint, we are a customer-focused community, dedicated to crafting solutions that redefine our industry. We have several products that provide essential services for businesses, such as credit, overdrafts, etc. We leverage artificial intelligence and data to make our decisions, but also have the technology and data-driven best practices used to support our businesses. Curious about what makes Moniepoint an incredible place to work? Check out posts on how we cultivate a culture of innovation, teamwork, and growth.
Job Summary We are seeking a Site Reliability Engineer (SRE) responsible for ensuring our systems run smoothly and efficiently while engineering solutions to improve visibility, eliminate repetitive tasks, and increase system resilience. The ideal candidate will balance real-time on-call responsibilities with strategic engineering work to achieve sustainable and scalable service reliability. Responsibilities Participate in on-call rotations to detect and triage service and reliability issues across all environments.
Act as the Incident Commander during major incidents: initiating war room or bridge calls, coordinating cross-functional teams, providing timely and clear status updates to all stakeholders. Create and maintain meaningful dashboards and alerts. Work with development teams to instrument their code to ensure visibility. Develop automation to eliminate manual and repetitive operational tasks (toil) related to reliability across both applications and infrastructure.
Implement and track Service Level Indicators (SLIs) and Service Level Objectives (SLOs) defined by the engineering leadership. Investigate and resolve customer complaints escalated beyond L1 and L2 support, especially those involving performance, reliability, or complex system behavior. Requirements Minimum of 3 years of experience supporting enterprise applications as an SRE or similar role with proficiency in writing code in Java, Go or Python Good understanding of distributed systems concepts, microservices architecture and software design patterns.