Developer II | Retail IT | Retail IT Operations

Allan Gray Proprietary Limited · Cape Town, Western Cape

Stop applying one at a time.

JobAlertsZA auto-applies to South African jobs like this one for you, overnight. Upload your CV once — we do the applying.

Start free — we apply for you →

Job Summary

  • We are looking for a detail-oriented developer with hands-on technical experience and a strong service mindset. The ideal candidate is comfortable in a fast-paced operations environment, responds calmly under pressure, and enjoys solving operational problems through automation, better tooling, and improved observability.
  • This is an opportunity to join Retail IT Operations — a team dedicated to improving stability. security, efficiency, observability, and resilience for Retail IT systems' infrastructure.
  • As a Developer, you will combine operational support with technical delivery. You will work independently on operational tasks, contribute to larger initiatives, and collaborate with feature/domain and infrastructure teams across IT.
  • The team is also investing heavily in observability maturity and resilience engineering so we are looking for someone who can grow with that direction.

Job Responsibilities Production and environment support

  • Provide production and non-production support for Retail applications and infrastructure, including incident response, alert triage and scheduled 24/7 standby for critical outages
  • Investigate issues, perform root-cause analysis, coordinate resolution across teams, and document outcomes and mitigations
  • Ensure development environments remain stable and available for Retail IT teams

Change, release, and resilience

  • Participate in change management and support release deployments for Retail IT teams
  • Contribute to resilience engineering — disaster recovery/AZ failover, capacity and availability improvements

Platform engineering and automation

  • Deliver infrastructure-as-code and automation solutions that reduce manual effort and improve consistency
  • Implement platform improvements across monitoring, alerting, observability, patching, vulnerability remediation, cost/stability, and environment engineering
  • Proactively reduce repeat support effort through better monitoring, alerts, automation, and process fixes
  • Adhere to data security and compliance with involvement and improvements in monthly database masking and restores

Engineering quality, security, and knowledge

  • Contribute to code reviews and follow Retail IT engineering standards
  • Apply security considerations in design, implementation, and operational change — including patching and vulnerability remediation
  • Document runbooks, SOPs, and knowledge-base articles; support onboarding of new team members and maintain shared operational documentation
  • Participate in team ceremonies, contribute to the backlog, and collaborate professionally with dev/feature teams, infrastructure, vendors and other stakeholders
  • Document runbooks, SOPs, and knowledge-base articles; support onboarding of new team members and maintain shared operational documentation
  • Participate in team ceremonies, contribute to the backlog, and collaborate professionally with dev/feature teams, infrastructure, vendors and other stakeholders

Experience

  • At least 2 years experience in IT operations, platform engineering, or software development with production support exposure
  • Practical experience with incident management and structured troubleshooting

Infrastructure and platforms

  • Working knowledge of Linux and Windows servers, networking fundamentals, and common Retail IT infrastructure
  • Experience with SQL Server queries and basic administration
  • Familiarity with AWS, Terraform, Ansible, load-balancing, Kubernetes, microservices and distributed application architectures
  • Advantageous: understanding of system integration patterns, including RESTful services and RabbitMQ

Observability and reliability

  • Familiarity with monitoring and observability tools
  • Experience with metrics, logging, and tracing concepts; practical use of tools such as Grafana, Prometheus, Opensearch, OpenTelemetry, Cloudwatch and PagerDuty
  • Interest in growing observability and SRE practices

Automation and delivery

  • Ability to write and maintain scripts and automation bash/powershell
  • Experience with CI/CD pipeline concepts
  • Experience delivering operational or platform changes through repeatable, reviewable workflows

Security

  • Awareness of secure configuration, patching, and vulnerability remediation in operational environments
  • Ability to consider security implications in infrastructure and automation work; familiarity with organisational security standards

Ways of working

  • Clear verbal and written communication with technical and non-technical stakeholders
  • Strong attention to detail and problem-solving skills
  • Methodical approach; ability to balance reactive support with proactive improvement work
  • Planning, organisational, and time management skills

Education

  • Relevant IT or Computer Science qualification diploma or degree
  • Strong academic performance

Deadline:26th August,2026

Auto-apply to this jobView original posting ↗