ZfG Operations Engineer

Zoom Video Communications
San Jose
Workplace: HybridFull timeUSD 98,900 - 228,700 annuallyFunction: DevOps, Cloud & InfrastructureSkills: ["Collaboration","Problem-solving","Independent execution","Accountability","Knowledge sharing"]

Develop and maintain reliable infrastructure that powers Zoom’s distributed systems at global scale. Build automation and monitoring frameworks, analyze performance metrics to remove bottlenecks, and drive incident response through coordination, root-cause analysis, and preventive improvements. Maintain runbooks and service-level objectives, mentor teammates on troubleshooting and optimization best practices, and participate in on-call rotations supporting high-availability production systems.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Zoom Video Communications
Zoom Video Communications
2 days ago

ZfG Operations Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 11 hours agoStatus: Live

Job Summary

Develop and maintain reliable infrastructure that powers Zoom’s distributed systems at global scale. Build automation and monitoring frameworks, analyze performance metrics to remove bottlenecks, and drive incident response through coordination, root-cause analysis, and preventive improvements. Maintain runbooks and service-level objectives, mentor teammates on troubleshooting and optimization best practices, and participate in on-call rotations supporting high-availability production systems.
Location: San Jose
Workplace: Hybrid
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure

Key Responsibilities

  • •Design and implement automation frameworks and monitoring solutions to enhance reliability, reduce manual interventions, and optimize performance in distributed production environments.
  • •Analyze system performance metrics to identify bottlenecks, recommend scalability improvements, and proactively prevent service degradation.
  • •Lead incident response by coordinating cross-functional teams, performing root cause analysis, and implementing preventive measures to reduce outage risk.
  • •Develop and maintain operational documentation, runbooks, and service level objectives to standardize procedures and improve team efficiency.
  • •Mentor team members on troubleshooting, system optimization, and infrastructure best practices while facilitating knowledge sharing across teams.

Pay and Benefits

Salary: USD 98,900 - 228,700 annually

Key Requirements

  • •Proficiency in at least one programming or scripting language (Python, Go, Bash, or similar) to build automation tools and infrastructure solutions.
  • •Knowledge of distributed systems architecture, including scalability patterns, fault tolerance mechanisms, and performance optimization principles.
  • •Experience using monitoring tools, observability platforms, and metrics collection systems to maintain production visibility.
  • •Ability to execute incident management, including response coordination, root cause analysis, and remediation planning.
  • •Equivalent practical experience in site reliability engineering, DevOps, or infrastructure operations, including contributing to on-call rotations for high-availability systems.
Skills:CollaborationProblem-solvingIndependent executionAccountabilityKnowledge sharing
Tech Stack:PythonGoBash

Company Brief

Zoom Video Communications
Provides video conferencing, chat, phone, webinar, and collaboration software for businesses, schools, and individuals. Its platform is widely used for remote meetings, virtual events, and hybrid work communication.
Industry: SaaS
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: San Jose, United States
Founded: 2011
WebsiteLinkedIn