Job Description:

This is a 12 – month contract with our client in the banking sector. The working model is hybrid – 3 days at the office.

We are seeking a Senior DevOps & Site Reliability Engineer for our client in the banking sector to design, build, and maintain highly available, scalable, secure, and automated cloud platforms for mission-critical applications. This role combines software engineering, platform engineering, cloud infrastructure, and observability to drive operational excellence, system resilience, and rapid deployment velocity across delivery teams.

Key Responsibilities

  • Platform Engineering & IaC: Design and build cloud-native infrastructure, automated provisioning workflows, and reusable deployment templates.
  • DevOps & DevSecOps: Build CI/CD pipelines with integrated automated testing and security controls to enable secure, zero-downtime releases and rollback capabilities.
  • Site Reliability & Observability: Establish SLIs, SLOs, and SLAs; reduce MTTD and MTTR through logging, metrics, and alerting; and lead incident response and root-cause analysis.
  • Cloud Operations & Cost Optimization: Manage high-availability and disaster recovery setups while continuously optimizing system performance, security, and cloud expenditure.
  • Leadership & Strategy: Provide technical leadership, mentor junior engineers, and champion engineering best practices across cross-functional delivery teams.

Minimum Requirements

  • Education: Preferred bachelor's degree in computer science, IT, Software Engineering, Information Systems, or an equivalent technical qualification.
  • Experience:
  • 8+ years of experience in software engineering, infrastructure, cloud, DevOps, or platform engineering.
  • 5+ years of hands-on DevOps engineering experience.
  • 3+ years in SRE or production operations managing large-scale, mission-critical enterprise systems.
  • CI/CD & Containers: Hands-on experience with Azure DevOps, GitHub, Jenkins, Docker, Kubernetes, and Helm.
  • Observability & Scripting: Strong knowledge of monitoring tools (Dynatrace, Grafana, Prometheus, Azure Monitor) and scripting/programming in Python, PowerShell, Bash, C#, or Java.

Working Place:

Johannesburg, South Africa