Site Reliability Engineer
South Africa
Workplace: RemoteFull timeFunction: DevOps, Cloud & InfrastructureExperience: 5+ yearsSkills: ["Incident response","Troubleshooting","Cross-functional coordination","Stakeholder communication","Automation mindset"]Own end-to-end service reliability by participating in on-call rotations, triaging incidents, and acting as Incident Commander during major outages. Build visibility with dashboards, alerts, and application instrumentation, then improve resilience by defining SLIs/SLOs and implementing automation to reduce operational toil. Investigate escalated customer issues, especially complex performance and reliability problems, while partnering with development teams across cloud and distributed systems environments.
Loading
Loading job details...
Preparing the role view and application actions.

