Senior Site Reliability Engineer, Compute
San Mateo
Workplace: OnsiteFull timeUSD 243,290 - 295,250 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 6+ yearsEducation: bachelorsSkills: ["Collaboration","Problem-solving","Planning","Curiosity","Data-driven thinking"]Own and operate the Infrastructure Compute cell infrastructure system and related layers like service discovery and secrets management. Build Roblox’s private cloud, productionize Kubernetes-based infrastructure, and drive reliability best practices across the Compute team. Create fault-tolerant, observable tooling and libraries, automate cluster lifecycle processes, and implement production guardrails using load testing, monitoring, and canarying to improve capacity and service reliability.
Loading
Loading job details...
Preparing the role view and application actions.

