Site Reliability Engineer, Compute Platform
San Jose
Full timeFunction: DevOps, Cloud & InfrastructureEducation: bachelorsSkills: ["Problem-solving","Critical thinking","Written communication","Verbal communication","Ownership"]Join a newly established Compute Platform SRE team that ensures reliability for TikTok’s major data warehouse products, services, and query engines. You’ll uphold SLAs, lead incident response and postmortems, and continuously improve performance by analyzing reliability signals. Work closely with product and development teams to embed reliability into the software lifecycle, automate provisioning and scaling, and plan capacity based on growth and upcoming initiatives.

