Developer II | Retail IT | Retail IT Operations
Allan Gray Proprietary Limited · Cape Town, Western Cape
Stop applying one at a time.
JobAlertsZA auto-applies to South African jobs like this one for you, overnight. Upload your CV once — we do the applying.
Start free — we apply for you →Job Summary
- We are looking for a detail-oriented developer with hands-on technical experience and a strong service mindset. The ideal candidate is comfortable in a fast-paced operations environment, responds calmly under pressure, and enjoys solving operational problems through automation, better tooling, and improved observability.
- This is an opportunity to join Retail IT Operations — a team dedicated to improving stability. security, efficiency, observability, and resilience for Retail IT systems' infrastructure.
- As a Developer, you will combine operational support with technical delivery. You will work independently on operational tasks, contribute to larger initiatives, and collaborate with feature/domain and infrastructure teams across IT.
- The team is also investing heavily in observability maturity and resilience engineering so we are looking for someone who can grow with that direction.
Job Responsibilities Production and environment support
- Provide production and non-production support for Retail applications and infrastructure, including incident response, alert triage and scheduled 24/7 standby for critical outages
- Investigate issues, perform root-cause analysis, coordinate resolution across teams, and document outcomes and mitigations
- Ensure development environments remain stable and available for Retail IT teams
Change, release, and resilience
- Participate in change management and support release deployments for Retail IT teams
- Contribute to resilience engineering — disaster recovery/AZ failover, capacity and availability improvements
Platform engineering and automation
- Deliver infrastructure-as-code and automation solutions that reduce manual effort and improve consistency
- Implement platform improvements across monitoring, alerting, observability, patching, vulnerability remediation, cost/stability, and environment engineering
- Proactively reduce repeat support effort through better monitoring, alerts, automation, and process fixes
- Adhere to data security and compliance with involvement and improvements in monthly database masking and restores
Engineering quality, security, and knowledge
- Contribute to code reviews and follow Retail IT engineering standards
- Apply security considerations in design, implementation, and operational change — including patching and vulnerability remediation
- Document runbooks, SOPs, and knowledge-base articles; support onboarding of new team members and maintain shared operational documentation
- Participate in team ceremonies, contribute to the backlog, and collaborate professionally with dev/feature teams, infrastructure, vendors and other stakeholders
- Document runbooks, SOPs, and knowledge-base articles; support onboarding of new team members and maintain shared operational documentation
- Participate in team ceremonies, contribute to the backlog, and collaborate professionally with dev/feature teams, infrastructure, vendors and other stakeholders
Experience
- At least 2 years experience in IT operations, platform engineering, or software development with production support exposure
- Practical experience with incident management and structured troubleshooting
Infrastructure and platforms
- Working knowledge of Linux and Windows servers, networking fundamentals, and common Retail IT infrastructure
- Experience with SQL Server queries and basic administration
- Familiarity with AWS, Terraform, Ansible, load-balancing, Kubernetes, microservices and distributed application architectures
- Advantageous: understanding of system integration patterns, including RESTful services and RabbitMQ
Observability and reliability
- Familiarity with monitoring and observability tools
- Experience with metrics, logging, and tracing concepts; practical use of tools such as Grafana, Prometheus, Opensearch, OpenTelemetry, Cloudwatch and PagerDuty
- Interest in growing observability and SRE practices
Automation and delivery
- Ability to write and maintain scripts and automation bash/powershell
- Experience with CI/CD pipeline concepts
- Experience delivering operational or platform changes through repeatable, reviewable workflows
Security
- Awareness of secure configuration, patching, and vulnerability remediation in operational environments
- Ability to consider security implications in infrastructure and automation work; familiarity with organisational security standards
Ways of working
- Clear verbal and written communication with technical and non-technical stakeholders
- Strong attention to detail and problem-solving skills
- Methodical approach; ability to balance reactive support with proactive improvement work
- Planning, organisational, and time management skills
Education
- Relevant IT or Computer Science qualification diploma or degree
- Strong academic performance
Deadline:26th August,2026