Site Reliability Engineer
Tennessee, Memphis, Mississippi
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureExperience: 5+ yearsEducation: bachelorsSkills: ["Incident leadership","Calm technical leadership","Problem-solving","Data-driven approach","Cross-functional collaboration"]Design and operate campus-level reliability systems by defining what is monitored, trusted, and alerted. Lead SEV-class incidents with technical command and NOC coordination, run blameless postmortems, and drive corrective actions to closure. Own monitoring signal quality and alert hygiene, set error budgets and availability objectives, and partner on cross-discipline projects spanning compute, network, storage, power, and cooling. Build playbooks, run game days, and support 24/7 operations.
Loading
Loading job details...
Preparing the role view and application actions.

