Lead Site Reliability Engineer, Platforms

Zoom Video Communications
San Jose
Full timeUSD 124,000 - 271,200 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 8+ yearsSkills: ["Communication","Incident management","Mentorship"]

Lead SRE for Zoom’s DevOps Platforms organization, owning critical infrastructure across AWS/OCI cloud, data center orchestration, security services, and the Zoom for Government (ZfG) environment. Define technical roadmaps, design and scale Kubernetes and infrastructure automation, and drive SRE best practices like infrastructure as code, monitoring, and incident management. Partner across teams, mentor engineers, and improve reliability and compliance-ready systems that underpin Zoom’s global services.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Zoom Video Communications
Zoom Video Communications
5 days ago

Lead Site Reliability Engineer, Platforms

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 18 hours agoStatus: Live

Job Summary

Lead SRE for Zoom’s DevOps Platforms organization, owning critical infrastructure across AWS/OCI cloud, data center orchestration, security services, and the Zoom for Government (ZfG) environment. Define technical roadmaps, design and scale Kubernetes and infrastructure automation, and drive SRE best practices like infrastructure as code, monitoring, and incident management. Partner across teams, mentor engineers, and improve reliability and compliance-ready systems that underpin Zoom’s global services.
Location: San Jose
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Design and scale DevOps platform services including Kubernetes infrastructure, cloud systems, and compliance-ready environments.
  • •Define technical roadmaps and architectural direction for infrastructure automation and security.
  • •Partner with service teams to deliver solutions that improve reliability and efficiency.
  • •Establish and advocate for SRE best practices including infrastructure as code, monitoring, and incident management.
  • •Mentor team members through design, implementation, and production deployment of complex systems.

Pay and Benefits

Salary: USD 124,000 - 271,200 annually

Key Requirements

  • •8+ years of SRE or DevOps experience building and operating production infrastructure at scale.
  • •Proficiency in a programming language beyond scripting (e.g., Python, Go, Java).
  • •Experience deploying and managing CI/CD pipelines using tools like Git, Jenkins, Argo CD, or JFrog.
  • •Experience operating cloud infrastructure on AWS or OCI using Terraform and Kubernetes.
  • •Observability experience with tools such as ELK, Prometheus, and Grafana, plus a degree in CS (or equivalent practical experience).
Experience:8+ yearsSREDevOpsProduction infrastructureCloud infrastructureKubernetesCI/CDObservabilitySecurity
Education:
Skills:CommunicationIncident managementMentorship
Languages:ChineseMandarin
Tech Stack:PythonGoJavaAWSOCIKubernetesGitJenkinsArgo CDJFrogTerraformCI/CDELKPrometheusGrafanaInfrastructure as codeMonitoringAutomation

Eligibility

Nationality:US National
Visa:Green Card

Company Brief

Zoom Video Communications
Provides video conferencing, chat, phone, webinar, and collaboration software for businesses, schools, and individuals. Its platform is widely used for remote meetings, virtual events, and hybrid work communication.
Industry: SaaS
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: San Jose, United States
Founded: 2011
WebsiteLinkedIn