Site Reliability Engineer, System - System Service Global
ByteDance
Singapore
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureEducation: bachelorsSkills: ["Self-motivated","Collaboration","Troubleshooting","Continuous improvement","Incident response"]Own reliability for ByteDance’s non-China data center infrastructure services, spanning core foundational components like DNS, NTP, DHCP, NAT, APT repositories, and Kerberos. Build high-availability deployment architectures with fault tolerance and disaster recovery, define SLOs/SLIs, and lead incident response with blameless post-mortems. Partner with network, security, and application teams while automating host management to reduce toil and improve operational efficiency.

