Site Reliability Engineer, Compute Platform
ByteDance
San Jose
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureEducation: bachelorsSkills: ["Problem-solving","Critical thinking","Written communication","Verbal communication","Collaboration"]Build and operate the reliability layer for major Big Data services and products, including data warehouse products and query engines. Own SLAs, respond to outages, and run incident management with troubleshooting and postmortems. Continuously optimize performance by analyzing reliability patterns, automating infrastructure provisioning and scaling, and collaborating with product and development teams. Forecast capacity and demand to support growth.

