Site Reliability Engineer - Data

ByteDance
San Jose
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureExperience: 3+ yearsEducation: bachelorsSkills: ["Problem solving","Communication"]

Join ByteDance’s Site Reliability Engineering organization to build and operate cloud-managed data infrastructure for Data AML, Data Center & Supply Chain, Data Infrastructure, and Data Architecture. You’ll improve the end-to-end service lifecycle, design monitoring and automation for service-oriented architecture, and develop scalable platform components using tools like Kubernetes, Redis, MySQL, and Flink. Handle incidents and drive blameless postmortems to continuously raise reliability and velocity.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
ByteDance
ByteDance
1 month ago

Site Reliability Engineer - Data

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 31 days agoStatus: Live
Reposted: similar role first listed 1 month ago

Job Summary

Join ByteDance’s Site Reliability Engineering organization to build and operate cloud-managed data infrastructure for Data AML, Data Center & Supply Chain, Data Infrastructure, and Data Architecture. You’ll improve the end-to-end service lifecycle, design monitoring and automation for service-oriented architecture, and develop scalable platform components using tools like Kubernetes, Redis, MySQL, and Flink. Handle incidents and drive blameless postmortems to continuously raise reliability and velocity.
Location: San Jose
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Participate in and improve the complete service lifecycle from inception and design through development, capacity planning, launch reviews, deployment, operation, and refinement.
  • •Design and implement software platforms and monitoring frameworks to govern service-oriented architecture (SOA) efficiently, automatically, and intelligently.
  • •Develop and manage components of cloud-managed data infrastructure using technologies such as Kubernetes, Redis, MySQL, Flink, and more.
  • •Establish automation-driven scaling mechanisms to improve reliability, efficiency, and velocity.
  • •Provide user support, manage incident responses, and run blameless postmortems to improve systems.

Key Requirements

  • •Bachelor’s degree in Computer Science or a related technical field.
  • •At least 3 years of experience programming in C, C++, Java, Python, Go, or Rust.
  • •Familiarity with Unix/Linux system internals, networking, and distributed systems.
  • •Preferred: experience with MySQL and Redis.
  • •Preferred: experience designing and analyzing large-scale distributed systems.
Experience:3+ years
Education:Bachelor's
Skills:Problem solvingCommunication
Tech Stack:CC++JavaPythonGoRustUnix/LinuxNetworkingDistributed systemsKubernetesRedisMySQLFlinkService-oriented architecture (SOA)DockerOpenStackHadoopSparkAutomationMonitoring frameworks

Company Brief

ByteDance
Develops consumer internet and content platforms, including TikTok and other apps for short-form video, news, and entertainment. It also builds advertising, commerce, and creator tools that connect audiences, brands, and publishers across global markets.
Industry: Digital Media
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Established Company
Headquarters: Beijing, China
Founded: 2012
WebsiteLinkedIn