Site Reliability Engineer - Data Infrastructure
San Jose
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureExperience: 2+ yearsEducation: bachelorsSkills: ["Problem-solving","Communication","Proactive attitude","Desire to learn"]Own the reliability, scalability, and efficiency of core data services powering TikTok products. This SRE role responds to production incidents via runbooks, performs change-controlled deployments and maintenance, and improves observability through dashboards, alert tuning, and instrumentation. You’ll automate repetitive operations with scripting and AI augmentation, and support daily operations and upkeep of data center and AI infrastructure for large-scale data processing.

