Moniepoint is Africa's all-in-one financial ecosystem, empowering businesses and their customers with seamless payment, banking, credit, and management tools. In 2023, we processed $182 billion and are Nigeria’s largest merchant acquirer. We are  recruiting to fill the position below:

 Job Title: Senior Site Reliability Engineer

Location: Remote Job Summary

We are seeking an experienced SRE to engineer the reliability of our highly distributed platform. You will combine deep knowledge of distributed systems with strong coding skills to define SLOs, lead incident response, and build automation and self-healing mechanisms into our systems. You will balance immediate operational stability with long-term strategic engineering to ensure our services scale linearly with our hyper-growth.


Responsibilities

Participate in on-call rotations as the primary technical lead. Act as the Incident Commander during major severity incidents: initiating war rooms, coordinating cross-functional teams, and providing clear status updates. Instrument code to expose high-cardinality metrics and distributed traces. Collaboratively define, measure, and defend Service Level Objectives (SLOs) and Error Budgets with product owners. Write high-quality, production-ready code (in Java, Go, or Python) to build internal tooling, automation platforms, and self-healing mechanisms that eliminate manual operator intervention. Partner with Product Engineering teams during the design phase to ensure new services are built with reliability, scalability, and observability patterns (circuit breakers, rate limiting, backpressure, fallback strategies) from day one. Analyze system performance and traffic patterns to model future capacity needs. Conduct load testing and chaos engineering experiments to verify system resilience under failure conditions.


Requirements

Minimum of 5 years of experience in SRE or Backend Engineering with a strong ability to write clean, performant, and tested code in Java, Go, Rust, or Python. Deep understanding of distributed systems architecture and design patterns. You possess a strong command of microservices fundamentals, event-driven architectures, and the underlying principles required to build systems that scale. Extensive experience with Google Cloud Platform (GCP) or similar cloud providers (AWS/Azure). You are proficient in running production workloads on Kubernetes (GKE/EKS) and troubleshooting cluster/infrastructure issues. Experience designing observability strategies using OpenTelemetry, Prometheus, New Relic, Datadog, or SigNoz to improve system visibility. Familiarity with operating and tuning production data stores (e.g., PostgreSQL, MySQL) and streaming platforms (e.g., Kafka, RabbitMQ) in a high-throughput environment.


Application Closing Date

Not Specified. Method of Application


Interested and qualified candidates should:

Click Apply Now button to apply

Would you like WhatsApp and Email notifications on more Senior Site Reliability Engineer (Remote)  Remote job offers? Subscribe to our Premium Job Alert for customised job alerts, access to our Database Subscription WhatsApp Group and profile recommendation to hiring employers.
SUBSCRIBE TO PREMIUM ALERT

Salary

0 - 0 NGN

Monthly based

Remote Job

Worldwide

Job Overview
Job Posted:
1 day ago
Job Expire:
1 month from now
Job Type
Intern
Job Role
Entry level role
Education
Any Level / PhD
Total Vacancies
Variable
Category
Administration / Secretarial

Share This Job:

Location

 Remote

Advance your career with BetaJob Certificates

Create account, take courses and earn verifiable certificates. Download and add certifications to your profile for better chance of getting hired.

Subscribe To Premium Alert To Get Hired Faster!

Get seen by employers With premium alert subscription, your profile is recommended to employers hiring matching criteria you indicate.
Access the best jobs for you Sign up for customised job alerts matching your experience, preferred industry, function and location.