SRE Engineer - Seattle
Plaud
Seattle
Workplace: HybridFull timeFunction: DevOps, Cloud & InfrastructureExperience: 5+ yearsSkills: ["Incident response","Reliability engineering","Communication"]Own production reliability for AI workloads by designing and operating highly available, scalable cloud-native systems. Build observability with metrics, logs, and tracing, automate reliability, and define SLOs/SLIs and error budgets. Lead incident response, run postmortems, and drive continuous reliability improvements while partnering with product and engineering teams on reliability-focused design.

