Staff Site Reliability Engineer

Zscaler
San Jose
Workplace: HybridFull timeUSD 119,000 - 170,000 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 5+ yearsSkills: ["Ownership","Problem-solving","Collaboration","Urgency","Scale","On-call","Automation","Monitoring","Communication"]

Join the Zscaler Cloud Operations team as a Staff Site Reliability Engineer to own and optimize the production data center services, ensuring availability, latency, and scalability for a global cloud platform. Design and deploy software and automation, build scalable monitoring, participate in on-call rotations, and collaborate with cross-functional teams to improve processes at scale.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Zscaler
Zscaler
6 months ago

Staff Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 23 hours agoStatus: Live

Job Summary

Join the Zscaler Cloud Operations team as a Staff Site Reliability Engineer to own and optimize the production data center services, ensuring availability, latency, and scalability for a global cloud platform. Design and deploy software and automation, build scalable monitoring, participate in on-call rotations, and collaborate with cross-functional teams to improve processes at scale.
Location: San Jose
Workplace: Hybrid
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure

Key Responsibilities

  • •Design, code, and deploy software solutions and automation while looking for opportunities to optimize the existing code-base for maintainability and reusability
  • •Create and deploy scalable monitoring systems and end-to-end solutions for a massively growing global infrastructure in collaboration with Software Engineering and Development teams
  • •Monitor applications and services within the environments, participate in on-call rotation, and implement strategies to prevent future occurrences of issues
  • •Resolve escalated issues and prevent recurring operational overhead by documenting and automating processes while deploying patches, upgrades, and administrative tools
  • •Collaborate with cross-functional teams to recommend integration strategies for platforms and applications to constantly improve and identify opportunities for process improvement

Pay and Benefits

Salary: USD 119,000 - 170,000 annually
Perks:Health InsuranceParental LeaveRemote Work

Key Requirements

  • •US Citizenship is required and 5+ years of industry experience in a 24/7 NOC or Cloud Operations environment
  • •Proficiency with programming languages such as Python or Bash
  • •Deep understanding of networking standard protocols including HTTP, DNS, TCP/IP, ICMP, and the OSI Model
  • •Hands-on experience with monitoring tools (e.g. Nagios, Grafana, Prometheus, etc.) and networking principles like Firewalls and Load Balancing
  • •Ability and flexibility to work after hours or weekends for application releases and deployments in a fast-paced environment
Experience:5+ yearsCloud operationsNOCCloud computingIT operations
Skills:OwnershipProblem-solvingCollaborationUrgencyScaleOn-callAutomationMonitoringCommunication
Languages:English
Tech Stack:PythonBashNagiosGrafanaPrometheusGoHTTPDNSTCP/IPOSI ModelOn-callAutomation

Company Brief

Zscaler
Provides cloud-native security platform delivering secure access, threat protection, and zero trust services to organizations, enabling secure internet and private application access without traditional network appliances.
Industry: Cybersecurity
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: San Jose, United States
Founded: 2007
WebsiteLinkedIn