Engineering

Site Reliability Engineer (sre) At Hampshire Heights Global Company Limited

Hampshire Heights Global Company Limited·Lagos, nigeria·Full Time·Onsite
EngineeringFull TimeOnsite

Interested applicant(s) should possess relevant qualifications and experience

Key skills

BA/BSc/HND
Apply Now
Send your CV along with a cover letter to[email protected]

Please use the job title as the subject line of your email.

At a glance

Company

Hampshire Heights Global Company Limited

Location

Lagos, nigeria

Employment

Full Time

Work style

Onsite

Experience

Valid until

Not specified

Created

July 28, 2026

More opportunities

Similar roles you might like

Site Reliability Engineer At Moniepoint Inc.

Moniepoint Inc.

All, nigeria

full-time

Job Summary We are seeking a Site Reliability Engineer (SRE) responsible for ensuring our systems run smoothly and efficiently while engineering solutions to improve visibility, eliminate repetitive tasks, and increase system resilience. The ideal candidate will balance real-time on-call responsibilities with strategic engineering work to achieve sustainable and scalable service reliability. Responsibilities Participate in on-call rotations to detect and triage service and reliability issues across all environments. Act as the Incident Commander during major incidents: initiating war room or bridge calls, coordinating cross-functional teams, providing timely and clear status updates to all stakeholders. Create and maintain meaningful dashboards and alerts. Work with development teams to instrument their code to ensure visibility. Develop automation to eliminate manual and repetitive operational tasks (toil) related to reliability across both applications and infrastructure. Implement and track Service Level Indicators (SLIs) and Service Level Objectives (SLOs) defined by the engineering leadership. Investigate and resolve customer complaints escalated beyond L1 and L2 support, especially those involving performance, reliability, or complex system behavior. Requirements Minimum of 4 years of experience supporting enterprise applications as an SRE or similar role with proficiency in writing code in Java, Go or Python Good understanding of distributed systems concepts, microservices architecture and software design patterns. Hands-on experience with Kubernetes. You have managed applications on a major cloud provider (GCP, AWS, or Azure), and can troubleshoot common container issues. Experience setting up dashboards in Grafana and using APM tools like Datadog, New Relic, Signoz. You have a Solid understanding of metrics, logs, and traces. Proficiency in SQL (e.g., PostgreSQL, MySQL). Ability to write complex queries to debug data issues and a basic understanding of database performance.

2 months ago

Site Reliability Engineer (sre) At Renmoney

Renmoney

Lagos, nigeria

full-time

The Role The Site Reliability Engineer (SRE) is responsible for ensuring the availability, reliability, scalability, and performance of business-critical applications and infrastructure. The role combines software engineering and operations expertise to automate processes, improve platform stability, and enhance system observability. What You Will Do Design, implement, and maintain highly available and scalable infrastructure. Monitor production systems and proactively identify performance bottlenecks. Manage incident response, root cause analysis (RCA), and problem management activities. Develop automation scripts and tools to improve operational efficiency. Implement and maintain CI/CD pipelines. Manage cloud infrastructure across AWS and hybrid environments. Configure and maintain observability platforms including monitoring, logging, and alerting solutions. Define and track SLIs, SLOs, and error budgets. Support application deployments and release management processes. Collaborate with Engineering, Security, Data, and Product teams to improve system reliability. Perform capacity planning and disaster recovery testing. Ensure infrastructure and systems comply with security and regulatory requirements. Requirements What You Bring Education Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field. Experience 4 - 7 years of experience in Site Reliability Engineering, DevOps, Cloud Engineering, or Infrastructure Operations. Experience supporting mission-critical financial services or fintech platforms is an advantage. Technical Skills Strong knowledge of AWS services (EC2, ECS/EKS, RDS, Lambda, VPC, IAM, CloudWatch). Experience with Infrastructure as Code (Terraform, CloudFormation). Knowledge of containerization technologies (Docker, Kubernetes). Experience with CI/CD tools (GitHub Actions, GitLab CI/CD, Jenkins, Azure DevOps). Experience with monitoring tools such as Datadog, Prometheus, Grafana, New Relic, or ELK Stack. Strong Linux administration skills. Experience with scripting languages (Python, Bash, PowerShell). Understanding of networking, DNS, load balancing, VPNs, and security controls. Preferred Certifications AWS Certified Solutions Architect. AWS SysOps Administrator. Kubernetes Certifications (CKA/CKAD). HashiCorp Terraform Associate. Key Competencies Problem-solving and analytical thinking. Incident management and troubleshooting. Automation mindset. Strong communication and collaboration. Attention to detail. This Role Is Ideal For You If You enjoy solving complex infrastructure and reliability challenges. You are passionate about automation and reducing operational overhead. You thrive in highly available, customer-facing environments where up-time matters. You enjoy working across Engineering, Security, Data, and Product teams to improve system performance. You are proactive and constantly seek opportunities to improve reliability, scalability, and efficiency. You May Not Enjoy This Role If You prefer manual processes over automation. You are uncomfortable responding to production incidents and troubleshooting critical issues. You prefer working in isolated environments with limited collaboration. You are not interested in continuous learning and evolving cloud technologies.

3 months ago

Site Reliability Engineer At Moniepoint Inc.

Moniepoint Inc.

nigeria

full-time

Job Summary We are seeking a Site Reliability Engineer (SRE) responsible for ensuring our systems run smoothly and efficiently while engineering solutions to improve visibility, eliminate repetitive tasks, and increase system resilience. The ideal candidate will balance real-time on-call responsibilities with strategic engineering work to achieve sustainable and scalable service reliability. Responsibilities Participate in on-call rotations to detect and triage service and reliability issues across all environments. Act as the Incident Commander during major incidents: initiating war room or bridge calls, coordinating cross-functional teams, providing timely and clear status updates to all stakeholders. Create and maintain meaningful dashboards and alerts. Work with development teams to instrument their code to ensure visibility. Develop automation to eliminate manual and repetitive operational tasks (toil) related to reliability across both applications and infrastructure. Implement and track Service Level Indicators (SLIs) and Service Level Objectives (SLOs) defined by the engineering leadership. Investigate and resolve customer complaints escalated beyond L1 and L2 support, especially those involving performance, reliability, or complex system behavior. Requirements Minimum of 3 years of experience supporting enterprise applications as an SRE or similar role with proficiency in writing code in Java, Go or Python Good understanding of distributed systems concepts, microservices architecture and software design patterns. Hands-on experience with Kubernetes. You have managed applications on a major cloud provider (GCP, AWS, or Azure), and can troubleshoot common container issues. Experience setting up dashboards in Grafana and using APM tools like Datadog, New Relic, Signoz. You have a Solid understanding of metrics, logs, and traces. Proficiency in SQL (e.g., PostgreSQL, MySQL). Ability to write complex queries to debug data issues and a basic understanding of database performance. What we can offer you Culture - We put our people first and prioritize the well-being of every team member. We've built a company where all opinions carry weight and where all voices are heard. We value and respect each other and always look out for one another. Above all, we are human. Learning - We have a learning and development-focused environment with an emphasis on knowledge sharing, training, and regular internal technical talks. Compensation - You'll receive an attractive salary, pension, health insurance, annual bonus, plus other benefits. What to expect in the hiring process: A preliminary phone call with the recruiter A technical interview with the Hiring Manager A behavioural and technical interview with a member of the Executive team. Moniepoint is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees and candidates.

3 months ago

Site Reliability Engineer At Cowrywise

Cowrywise

Lagos, nigeria

full-time

We're looking for a Site Reliability Engineer (SRE) to help build, maintain, and scale the infrastructure powering Cowrywise.You'll work closely with our engineering team to improve reliability, observability, security, and deployment processes across our systems. Our infrastructure team specializes across four areas: Cloud, Databases, Platform, and Observability. We run primarily on AWS with some workloads on GCP. For this role, we're particularly interested in someone who can raise the bar on observability, helping us detect issues faster and resolve them with confidence. What you'll do Generally, members of the infrastructure team are able to do the following Design, maintain, and improve cloud infrastructure and internal platforms Improve system reliability, scalability, and performance across services Build and maintain CI/CD pipelines and deployment workflows Implement monitoring, logging, alerting, and observability systems Respond to incidents, troubleshoot production issues, and lead root cause analysis Automate operational tasks and infrastructure provisioning Work with engineering teams to improve service architecture and operational readiness Improve security posture, access controls, and infrastructure best practices Manage containerized workloads and orchestration platforms Maintain disaster recovery, backup, and high availability strategies What we're looking for Required 4+ years of experience in an SRE, DevOps, or Platform Engineering role running production systems Strong hands-on experience with AWS (compute, networking, IAM, storage, managed services) Deep expertise in observability designing meaningful metrics, dashboards, alerts, and SLOs that actually catch problems before users do Hands-on experience with New Relic, Grafana, and Prometheus (or equivalent tooling) A track record of reducing MTTD and MTTR through better instrumentation, alerting, and incident response practices Proficiency with Docker and containerized workflows Solid scripting and automation skills (Python, Bash, Go, or similar) Experience with infrastructure-as-code (Terraform, Pulumi, or CloudFormation) Strong Linux fundamentals and networking knowledge Experience building and maintaining CI/CD pipelines Comfort leading incident response and writing clear post-mortems Nice to have Experience operating Kubernetes in production Exposure to GCP or multi-cloud environments Background in one of our specialization areas: Databases (Postgres, MySQL, Redis), Platform engineering, or Cloud architecture Security-focused experience (IAM hardening, secrets management, compliance frameworks) Experience in fintech or other regulated, high-availability environments The people who succeed on this team People who are proactive and take ownership Engineers who automate before repeating manual work People who stay calm and methodical during incidents Engineers who care about clean systems and operational excellence Strong collaborators who work well across teams Curious builders who enjoy learning and improving systems continuously

3 months ago

Site Reliability Engineer (sre) At Hampshire Heights Global Company Limited
Lagos, nigeria
Apply