Engineering

Site Reliability Engineer

Nomba (Formerly Kudi)·Lagos, nigeria·Full Time·Onsite
EngineeringFull TimeOnsite
-

About the role

  • Implement and maintain highly available, scalable, and secure production systems, emphasising automation and Infrastructure as Code (IaC) principles.
  • Collaborate with software development teams to influence the architecture and design of applications for better scalability, reliability, and performance.
  • Develop and maintain monitoring, alerting, and logging solutions to proactively detect and resolve system issues.
  • Respond to incidents and outages, conducting root cause analysis, and implementing preventative measures to minimize future occurrences.
  • Participate in on-call rotations and provide timely response to critical incidents.
  • Continuously improve system performance through performance tuning, capacity planning, and load testing.
  • Implement security best practices, ensuring that systems are compliant with industry standards and regulations.
  • Automate routine operational tasks using scripting and programming languages.
  • Work with cross-functional teams to define and document operational procedures and runbooks.
  • Contribute to the improvement of the CI/CD pipelines to ensure seamless deployments.
  • Keep abreast of industry trends, emerging technologies, and best practices in SRE and cloud infrastructure management.

Requirements

  • Bachelor's Degree in Computer Science, Engineering, or a related field (or equivalent practical experience).
  • Proven experience as a Site Reliability Engineer, DevOps Engineer, or a similar role managing large-scale, highly available production systems.
  • Solid experience with cloud platforms (e.g., AWS, Azure, GCP), including proficiency in provisioning and managing resources.
  • Strong understanding of Linux/Unix systems and command-line utilities.
  • Proficiency in at least one programming or scripting language (e.g., Python, Ruby, Bash, PowerShell).
  • Experience with Infrastructure as Code (IaC) tools like Terraform or CloudFormation.
  • Knowledge of containerization and orchestration technologies (e.g., Docker, Kubernetes).
  • Familiarity with monitoring tools and concepts (e.g., Prometheus, Grafana, ELK stack).
  • Understanding of networking protocols, load balancing, and firewalls.
  • Strong problem-solving and troubleshooting skills, with a focus on root cause analysis.
  • Excellent communication and collaboration skills to work effectively with cross-functional teams.

Nice to have:

  • Relevant certifications in SRE, DevOps, or cloud technologies.
  • Experience with databases and data management (e.g., SQL, NoSQL, caching systems).
  • Knowledge of configuration management tools (e.g., Ansible, Puppet, Chef).
  • Understanding of Agile methodologies and experience in Agile/Scrum environments.
  • Familiarity with security practices and compliance frameworks.

Key skills

BA/BSc/HNDProfessional Certificate

At a glance

Company

Nomba (Formerly Kudi)

Location

Lagos, nigeria

Employment

Full Time

Experience

Valid until

Not specified

Created

October 8, 2026

More opportunities

Similar roles you might like

Devops Manager

Seven Up Bottling Company

Abuja, nigeria

full-time

Role Description The DevOps Manager is a full-time, on-site role based in Abuja, responsible for leading and coordinating DevOps activities across software and infrastructure teams. This role oversees the design, implementation, and maintenance of CI/CD pipelines, ensuring reliable integration, automated testing, and smooth deployment of applications and services. The DevOps Manager will manage infrastructure provisioning and configuration, monitor system performance, and establish best practices for security, scalability, and resilience. Day-to-day tasks include collaborating with software development teams, guiding test automation and integration processes, resolving deployment issues, and mentoring DevOps engineers to improve delivery speed and system stability. The role also involves continuous improvement of tools, workflows, and documentation to support efficient and compliant operations. Qualifications Strong experience with Software Development and Integration, with the ability to collaborate closely with engineering teams on build, release, and deployment processes. Hands-on expertise in Continuous Integration practices and tools, including configuring and optimizing CI/CD pipelines for multiple applications and environments. Proficiency in Test Automation, including designing and integrating automated tests into deployment workflows to ensure reliability and quality. Solid background in Infrastructure management, such as server provisioning, cloud or on-premise environments, configuration management, monitoring, and incident response. Proven experience in a DevOps or Site Reliability Engineering leadership role, managing teams and cross-functional stakeholders. Knowledge of containerization, orchestration, and modern deployment frameworks (e.g., Docker, Kubernetes, infrastructure-as-code tools). Strong problem-solving skills, attention to detail, and ability to operate in a high-availability, production-focused environment. Excellent communication and collaboration skills, with the ability to document processes and provide clear technical guidance. Bachelor's degree in Computer Science, Engineering, Information Technology, or a related field; relevant professional certifications are an advantage.

7 days ago

Senior Devops Engineer

Paystack

Lagos, nigeria

full-time

About the Senior DevOps Engineer role As a Senior DevOps Engineer at Paystack, you will play a pivotal role in designing, implementing, and maintaining our cloud infrastructure, CI/CD pipelines, and automation frameworks. You will collaborate closely with cross-functional teams to optimise our development and deployment processes, improve system reliability, and drive continuous improvement initiatives. The ideal candidate will have a strong background in DevOps practices, cloud technologies, automation tools, database management, networking, and a passion for solving complex technical challenges. The selected engineer will join a team of DevOps engineers and report directly to the Lead DevOps Engineer. You will work closely with other members of the department including members of the larger infrastructure and security teams and quite heavily with other product engineering teams. We'll trust you to Own projects end to end, often across teams. Design technical solutions that are scalable and maintainable Design systems for high availability and fault tolerance. Build end to end automation pipelines. Tackle ambiguous, high risk problems proactively. Introduce new ideas and simplifies complex systems. Design secure systems, manages secrets and implements zero trust models. Participate in architectural evolution and software development activities. Help debug and diagnose issues that may arise in an intertwined distributed system environment. Own CI/CD architecture, minimises downtime and ensures security Lead major incident response and postmortems. You'll thrive in this role if you have 5+ years of experience in a DevOps or Site Reliability Engineering (SRE) role. Deep Expertise with a variety of AWS services Experience with IAC and tools like Terraform, designing reusable modules and enforces standards Real world experience with Kubernetes at scale Experience writing scripts and automation using Python, Node.js, Bash, Golang Experience with networking concepts and protocols, including TCP/IP, DNS, HTTP/S, VPNs, etc. Experience building tooling to reduce occurrences of errors and improve engineering productivity Experience with CI/CD tooling like ArgoCD, Github Actions, etc. Experience working with and monitoring and alerting tools like Prometheus and Grafana. Ability to clearly articulate, implement and document design decisions Bonus points if you have: Bachelor's degree in Computer Science, Engineering, or a related field might be beneficial but not required Experience in the fintech industry or with payment processing systems Knowledge of cybersecurity best practices and compliance requirements (e.g., PCI DSS) Experience with machine learning and data analytics. Benefits Competitive compensation package and benefits 13th month bonus TSG Equity compensation Full medical coverage Wellbeing stipend Generous leave and sabbatical policies Hybrid working environment Smart, kind colleagues who are invested in your growth.

a month ago

Devops & Cloud Engineer

Corona Management Systems

Abuja, nigeria

full-time

Job Description We're looking for a skilled DevOps and Cloud Engineer to help design, build, and maintain the infrastructure that powers our products. You'll work at the intersection of development and operations, automating processes, optimizing cloud environments, and ensuring our systems are reliable, secure, and scalable. This is a great opportunity for someone who enjoys solving complex infrastructure challenges and wants to have a direct impact on how we build and ship software. What You'll Do Design, implement, and maintain scalable, secure cloud infrastructure (AWS/Azure/GCP) Build and improve CI/CD pipelines to streamline deployment processes Automate infrastructure provisioning using Infrastructure as Code (Terraform, CloudFormation, or similar) Monitor system performance, availability, and reliability; respond to incidents and drive root-cause analysis Implement and maintain containerization and orchestration solutions (Docker, Kubernetes) Collaborate with software engineers to optimize application deployment and performance Strengthen security practices across infrastructure, including access controls, secrets management, and vulnerability remediation Manage configuration and environment consistency across development, staging, and production Contribute to on-call rotation and support incident response Document infrastructure, processes, and runbooks for team-wide visibility Requirements 3+ years of experience in DevOps, Site Reliability Engineering, or Cloud Engineering roles Hands-on experience with a major cloud provider (AWS, Azure, or GCP) Proficiency with Infrastructure as Code tools (Terraform, CloudFormation, Pulumi, etc.) Experience with containerization and orchestration (Docker, Kubernetes) Strong scripting skills (Python, Bash, or Go) Familiarity with CI/CD tools (Jenkins, GitLab CI, GitHub Actions, CircleCI, etc.) Solid understanding of networking, security, and system administration fundamentals Experience with monitoring and logging tools (Prometheus, Grafana, Datadog, ELK, etc.) Nice to Have: Relevant cloud certifications (AWS Certified Solutions Architect, Azure Administrator, etc.) Experience with GitOps workflows (ArgoCD, Flux) Familiarity with configuration management tools (Ansible, Chef, Puppet) Experience working in regulated or high-compliance environments Background in cost optimization and cloud governance

2 months ago

Engineering Manager (site Reliability)

Moniepoint Inc.

Lagos, nigeria

full-time

About the role As an Engineering Manager, you will drive the successful delivery and execution of projects within your teams. You will manage end-to-end technical planning, ensuring that product requirements are translated into actionable tasks, while orchestrating collaboration between various stakeholders, including engineers, product managers, QA, and UX. This role requires a deep understanding of software design and development and the ability to plan, execute, and deliver product features in a timely and predictable manner. You will also be responsible for maintaining high technical standards, managing team bandwidth, and ensuring project milestones are met with efficiency and accuracy. What you'll get to do Own delivery and execution within the Site Reliability Engineering tooling team. Evaluate product requirements for feasibility and ensure they align with the existing product architecture, translating them into EPICs and technical stories. Work closely with Product Managers and Engineers to refine and groom tasks. Plan and organize sprints with clearly defined goals, using project planning tools to establish timelines, and delivery milestones, and identify task dependencies early. Foster engineering processes that promote seamless collaboration and teamwork. Track team velocity to ensure resources are effectively allocated, balancing bandwidth with task demands. Coordinate alignment and manage dependencies across multiple stakeholders to prevent bottlenecks and ensure smooth execution. Contribute to critical projects by ensuring appropriate design patterns and coding techniques are applied. Remain hands-on, participating in code reviews to uphold high-quality standards. Ensure monitoring and observability are in place for all owned services, meeting defined SLIs/SLOs. Partner with Product Managers to track and publish post-deployment product metrics, ensuring transparency with key stakeholders. To succeed in this role, you should have BSc in Computer Science, Engineering, or a related field Minimum of 8 years of experience as a Software Developer, Software Engineer, or similar role. 5+ years of Node JS / python experience. A minimum of 3 years of leadership experience is a must. Strong understanding of agile methodologies, sprint planning, and backlog management. Expertise in breaking down complex product requirements into structured EPICs, Stories, and Tasks. Solid experience with backend technologies. Experience with frontend is a plus. Knowledge of project planning tools for visualizing and tracking delivery timelines. Familiarity with engineering metrics and monitoring tools to assess team performance and product health. Capability to debug complex technical issues during incidents to identify solutions and run blameless RCA sessions. Understanding of deployment pipelines, continuous integration (CI), continuous deployment (CD), and their corresponding metrics. Ability to drive alignment across diverse technical and non-technical stakeholders. Exceptional ability to manage dependencies, mitigate risks, and communicate clearly with stakeholders. Proven track record of improving team velocity and fostering efficient delivery. Generic Skills: Problem-solving: Ability to assess complex problems, find solutions, and make sound decisions. Communication: Strong written and verbal communication skills, including technical documentation and stakeholder reporting. Adaptability: Able to thrive in a fast-paced, changing environment, adjusting strategies as needed. Attention to Detail: Meticulous in documenting technical requirements and ensuring all aspects of a project are accounted for. Supervisory skills: Team Management: Experience in managing and mentoring engineers, ensuring team growth and performance. Resource Allocation: Ability to assess bandwidth and manage resource distribution to optimize team performance. Feedback: Conduct regular performance reviews, providing constructive feedback and fostering a growth-oriented environment. Stakeholder Management: Lead project status reviews, manage expectations, and ensure smooth communication between teams and leadership. What we can offer you Culture - We put our people first and prioritize the well-being of every team member. We've built a company where all opinions carry weight and where all voices are heard. We value and respect each other and always look out for one another. Above all, we are human. Learning - We have a learning and development-focused environment with an emphasis on knowledge sharing, training, and regular internal technical talks. Compensation - You'll receive an attractive salary, pension, health insurance, paid leave, plus other benefits.

2 months ago

Devops & Cloud Engineer

Corona Management Systems

Lagos, nigeria

full-time

About the Role We're looking for a skilled DevOps and Cloud Engineer to help design, build, and maintain the infrastructure that powers our products. You'll work at the intersection of development and operations, automating processes, optimizing cloud environments, and ensuring our systems are reliable, secure, and scalable. This is a great opportunity for someone who enjoys solving complex infrastructure challenges and wants to have a direct impact on how we build and ship software. What You'll Do Design, implement, and maintain scalable, secure cloud infrastructure (AWS / Azure / GCP). Build and improve CI / CD pipelines to streamline deployment processes. Automate infrastructure provisioning using Infrastructure as Code (Terraform, CloudFormation, or similar). Monitor system performance, availability, and reliability; respond to incidents and drive root-cause analysis. Implement and maintain containerization and orchestration solutions (Docker, Kubernetes). Collaborate with software engineers to optimize application deployment and performance. Strengthen security practices across infrastructure, including access controls, secrets management, and vulnerability remediation. Manage configuration and environment consistency across development, staging, and production. Contribute to on-call rotation and support incident response. Document infrastructure, processes, and runbooks for team-wide visibility. What We're Looking For Required: 3+ years of experience in DevOps, Site Reliability Engineering, or Cloud Engineering roles. Hands-on experience with a major cloud provider (AWS, Azure, or GCP). Proficiency with Infrastructure as Code tools (Terraform, CloudFormation, Pulumi, etc.). Experience with containerization and orchestration (Docker, Kubernetes) Strong scripting skills (Python, Bash, or Go). Familiarity with CI/CD tools (Jenkins, GitLab CI, GitHub Actions, CircleCI, etc.). Solid understanding of networking, security, and system administration fundamentals. Experience with monitoring and logging tools (Prometheus, Grafana, Datadog, ELK, etc.). Nice to Have: Relevant cloud certifications (AWS Certified Solutions Architect, Azure Administrator, etc.). Experience with GitOps workflows (ArgoCD, Flux). Familiarity with configuration management tools (Ansible, Chef, Puppet). Experience working in regulated or high-compliance environments. Background in cost optimization and cloud governance.

2 months ago

Devops Engineer

Kredete

Lagos, nigeria

full-time

Role Overview We are seeking a DevOps Engineer to architect, build, and maintain the robust and scalable infrastructure that powers our platform. In this pivotal role, you will bridge the gap between development and operations, automating processes, optimizing our CI/CD pipelines, and ensuring the reliability, security, and performance of our systems at scale. What You Will Do Design & Build Infrastructure: Architect, implement, and manage our cloud infrastructure (AWS & GCP) using Infrastructure as Code (IaC) tools like Terraform or CloudFormation. Master CI/CD Pipelines: Own, optimize, and secure the end-to-end CI/CD pipeline to enable rapid, reliable, and automated software deployments. Containerize & Orchestrate: Design and maintain containerized environments using Docker and orchestrate them with Kubernetes for scalable and efficient application deployment. Monitor & Optimize Performance: Implement and manage comprehensive monitoring, logging, and alerting solutions (e.g., Prometheus, Grafana, ELK Stack) to ensure system health and proactively resolve issues. Ensure Security & Compliance: Embed security best practices into the development lifecycle (DevSecOps), managing secrets, performing vulnerability scans, and ensuring infrastructure compliance. Collaborate & Automate: Work closely with development teams to foster a DevOps culture, identify bottlenecks, and automate everything from infrastructure provisioning to operational workflows. Drive Reliability: Practice incident management and lead post-mortem analyses to create a culture of continuous improvement and system resilience. Who You Are You hold a Bachelor's degree in Computer Science, Engineering, or a related field. You have 5-7 years of proven experience in a DevOps, Site Reliability Engineering, or similar role. You possess deep, hands-on expertise with: Cloud Platforms: Advanced knowledge of AWS (EC2, S3, RDS, IAM, VPC) and GCP (Experience with Azure/Azure DevOps will be a plus). Infrastructure as Code (IaC): Mastery of Terraform or CloudFormation. CI/CD Tools: Expert-level experience in designing and managing pipelines with GitHub Actions, ArgoCD (Experience with Jenkins, GitLab CI or Azure DevOps will be a plus). Containerization & Orchestration: Strong proficiency with Docker and Kubernetes, including cluster management and service deployment. Scripting & Automation: Fluency in Shell scripting and at least one language like Python or Go. Monitoring & Logging: Proven experience with tools like Prometheus, Grafana, Datadog, or the ELK Stack. Experience with Helm charts and the development/management of GitHub applications. You have a solid understanding of networking, security principles, and Linux system administration. You are an excellent problem-solver with a systematic approach to troubleshooting complex distributed systems. You have strong communication skills and thrive in a collaborative, fast-paced environment. Nice To Have Relevant professional certifications such as AWS & GCP Certified DevOps Engineer - Professional, CKAD (Kubernetes), or HashiCorp Terraform Associate. Experience with configuration management tools like Ansible, Puppet, or Chef. Knowledge of database administration and performance tuning for PostgreSQL, MySQL, or MongoDB. Experience in a microservices architecture and managing service mesh technologies (e.g., Istio, Linkerd). Background in a software development role, providing a strong understanding of the SDLC.

3 months ago