Senior Site Reliability Engineer
Kuala Lumpur
Workplace: HybridFull timeFunction: DevOps, Cloud & InfrastructureExperience: 5+ yearsSkills: ["Troubleshooting","Independent execution","Incident response","Operational improvement","Blameless postmortems"]Own reliability by operating SLI/SLO/error-budget targets, running end-to-end incident response, and driving blameless postmortems with tracked follow-through. Improve observability by extending instrumentation and raising dashboard/alert quality. Build safe release processes with canary and automated rollback, reduce toil through automation, and manage infrastructure as code. Lead production readiness reviews, performance/capacity planning, and infrastructure security. Operate the MCP gateway and apply AI-assisted tooling to operational workflows.
Loading
Loading job details...
Preparing the role view and application actions.

