Operations Engineer (Senior) - DevOps & Cloud

Abalobi Solutions · Pretoria, Gauteng

Stop applying one at a time.

JobAlertsZA auto-applies to South African jobs like this one for you, overnight. Upload your CV once — we do the applying.

Start free — we apply for you →

Introduction

An exciting opportunity is available for an experienced Senior Operations Engineer to join a highly technical, internationally integrated IT environment.

The successful candidate will play a key role in ensuring the reliability, availability, security and operational stability of enterprise applications and infrastructure across cloud and on-premises environments.

This position requires a strong combination of DevOps engineering, infrastructure automation, cloud operations, CI/CD, Infrastructure as Code, observability, incident management and IT Service Management .

Duties & Responsibilities

  • Lead configuration management and infrastructure automation initiatives.
  • Design, implement and maintain Infrastructure as Code (IaC) and automation solutions.
  • Build, maintain and optimise CI/CD pipelines .
  • Support and improve cloud and on-premises production environments.
  • Drive effective change, release and transition management .
  • Ensure configuration, release and change governance standards are maintained.
  • Act as a senior escalation point for complex production incidents.
  • Lead troubleshooting, Root Cause Analysis (RCA) and post-incident improvement activities.
  • Implement and enhance monitoring, logging and observability capabilities.
  • Collaborate with infrastructure and development teams to ensure solutions are designed for operational reliability and supportability.
  • Develop and maintain runbooks, operational procedures and technical documentation .
  • Monitor operational KPIs, service quality, availability and reliability.
  • Support IT Service Continuity , resilience and disaster-recovery requirements.
  • Facilitate technical onboarding, knowledge transfer and training.
  • Mentor and support junior Operations/DevOps engineers.
  • Drive continuous improvement and automation to reduce manual operational effort.
  • Ensure infrastructure and applications comply with security, lifecycle and governance requirements.
  • Maintain effective communication with technical and business stakeholders.
  • Support SLA/OLA and availability requirements.
  • Collaborate with geographically distributed and international DevOps teams.

Desired Experience & Qualification

ssential Technical Skills

Applicants should have strong practical experience in most of the following:

  • Configuration Management: Ansible, Puppet, Chef and/or Salt
  • Scripting & Automation: Python, Bash and/or PowerShell
  • CI/CD: Jenkins, GitLab CI, GitHub Actions or equivalent
  • Infrastructure as Code: Terraform, AWS CloudFormation or equivalent
  • Cloud Platforms: AWS, Azure or equivalent
  • Monitoring & Observability: Prometheus, Grafana, ELK/EFK or comparable enterprise monitoring solutions
  • Linux and/or Windows infrastructure administration
  • Production troubleshooting and Root Cause Analysis
  • Change, Release and Transition Management
  • ITSM/ITIL processes and operational governance
  • Git/version-control environments

Advantageous Skills

Experience in the following will be beneficial

  • Docker and Kubernetes
  • DevOps and Site Reliability Engineering (SRE) practices
  • SLIs, SLOs and error budgets
  • Middleware and enterprise platform technologies
  • Infrastructure security, hardening and lifecycle management
  • Data-centre infrastructure, networking, storage and servers
  • Highly regulated enterprise environments, particularly financial services or automotive
  • Deployment and testing automation
  • Data pipelines and/or ML deployment environments
  • Communities of Practice or Centres of Excellence
  • Mentoring/coaching junior technical professionals
  • German language capability

Qualifications & Experience

  • Relevant IT degree, diploma or equivalent qualification
  • Approximately 6–10 years' broad IT experience
  • At least 3–5 years' experience in IT Operations, DevOps, SRE, Cloud Operations or a closely related environment
  • Proven experience supporting complex production environments
  • Demonstrable experience with transition, change and release management
  • ITIL Foundation and Service Transition experience/qualification or equivalent is advantageous/highly preferred
  • Strong troubleshooting and analytical capability
  • Excellent stakeholder communication skills
  • Demonstrated ability to operate effectively within multidisciplinary technical teams
Auto-apply to this jobView original posting ↗