Team Lead, Site Reliability Engineering
Lagos, Nigeria · Full Time
Be the first to apply
- Experience
- 6+ yrs
- Salary
- —
- Openings
- 1
- Posted
- 3 days ago
- Work mode
- In office
- Resume
- Required to apply
Where you'll work
Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.
Job description
About Moniepoint Group
Moniepoint Inc. is a leading financial platform in Africa serving over 20 million businesses and individuals by providing seamless payment, banking, credit, cross-border, and business management solutions monthly. As Nigeria’s top merchant acquirer, the company supports the majority of the country’s POS transactions, processing digital payments exceeding $250 billion annually through its subsidiaries.
Role Overview
We are looking for a Team Lead for Site Reliability Engineering to oversee a group of engineers tasked with maintaining the reliability of our complex, distributed financial platform. This role involves high-level reliability architecture design, mentoring engineers, setting strategic technical direction, and fostering an SRE culture. You will balance technical leadership and hands-on responsibilities to support scalable growth during our rapid expansion.
Key Responsibilities
- Establish the technical vision and direction for the SRE squad, including architecting self-healing systems and setting reliability standards such as Production Readiness Reviews.
- Develop and enforce comprehensive system monitoring standards, guiding engineering teams in thorough code instrumentation (logging, tracing, metrics) and managing alerts to be actionable and minimize false alarms.
- Lead, mentor, and develop both senior and junior SRE team members by conducting code reviews, hosting technical workshops, and promoting engineering excellence.
- Serve as the primary escalation contact for major incidents, improving incident management processes and ensuring root cause analyses translate into effective engineering solutions.
- Collaborate with Engineering Managers and Product Leads to define Service Level Objectives aligned with organizational goals.
Candidate Requirements
- At least 6 years of experience in Site Reliability Engineering or Backend Engineering roles, with a minimum of 2 years in leadership or senior-level mentoring positions.
- Expert proficiency in at least one programming language among Java, Go, Rust, or Python, and capable of setting high standards for code quality.
- In-depth knowledge of distributed system design, including scalable architecture planning, troubleshooting complex microservices, and communicating architectural decisions effectively.
- Extensive experience with Google Cloud Platform or AWS, particularly managing and debugging Kubernetes (GKE) environments in production at scale.
- Proven ability to design and implement observability frameworks, covering custom instrumentation, monitoring tools, and alerting systems for large teams.
- Excellent communication skills, especially in managing critical incidents and maintaining composure to resolve high-pressure situations calmly.
What We Offer
- A respectful, inclusive workplace that prioritizes employee well-being and values all perspectives.
- Opportunities for continuous learning and professional development, including knowledge sharing, trainings, and technical talks.
- Attractive remuneration package consisting of competitive salary, pension, health insurance, annual bonus, and other benefits.
Hiring Process
- Initial conversation with a recruiter.
- Technical interview with the hiring manager.
- Behavioral and technical interview with an executive team member.
Equal Opportunity Employer
Moniepoint is committed to diversity and equality in the workplace and fosters an inclusive environment for all candidates and employees.