Major Incident Manager
Singapore
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureSkills: ["Change management","Infrastructure management","System support","Configuration management","Incident response"]Drive major incident operations and engineering improvements by using and building automation to monitor and observe production infrastructure across on-prem and cloud. Own system health monitoring, alert triaging, SLO/SLA-style error budget prioritization, and incident response through engineering actions from post-mortems. Strengthen security and observability through DevSecOps pipeline maintenance, infrastructure/policy as code, and continuous toil-elimination to improve reliability and MTTR.
Loading
Loading job details...
Preparing the role view and application actions.

