Manager, Site Reliability Engineering

Okta
San Francisco
Workplace: HybridFull timeUSD 204,000 - 306,000 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 3+ yearsSkills: ["Leadership","People management","Verbal communication","Written communication","Interpersonal skills"]

Lead Site Reliability Engineering efforts to scale Okta’s IDaaS platform—hosted on AWS across multiple availability zones and regions with 99.999 availability. Manage SREs and partner with architects to drive microservices reliability, DevOps maturity, and self-service automation. Build and improve CI/CD, observability platforms (Grafana, Splunk, APM), and self-healing patterns, while managing infrastructure-as-code (Terraform) and SDLC process improvements.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Okta
Okta
2 months ago

Manager, Site Reliability Engineering

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 14 hours agoStatus: Live
Reposted: similar role first listed 3 months ago

Job Summary

Lead Site Reliability Engineering efforts to scale Okta’s IDaaS platform—hosted on AWS across multiple availability zones and regions with 99.999 availability. Manage SREs and partner with architects to drive microservices reliability, DevOps maturity, and self-service automation. Build and improve CI/CD, observability platforms (Grafana, Splunk, APM), and self-healing patterns, while managing infrastructure-as-code (Terraform) and SDLC process improvements.
Location: San Francisco
Workplace: Hybrid
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Manager level

Key Responsibilities

  • •Manage a team of SREs supporting workloads for the IDaaS platform.
  • •Drive the microservice journey, DevOps maturity, and workload reliability with architects and teams across the organization.
  • •Accelerate SRE and product engineering velocity by building tooling, self-service capabilities, and self-healing patterns.
  • •Lead, mentor, and grow engineers and managers across platform, infrastructure, and shared services domains.
  • •Improve SDLC processes for cloud infrastructure as code, including CI/CD maturity and change/release management; manage project delivery within constraints.

Pay and Benefits

Salary: USD 204,000 - 306,000 annually
Perks:Health InsuranceDentalVision401kFlexible SpendingPaid LeaveParental Leave

Key Requirements

  • •3+ years of experience in technical leadership and people management.
  • •Extensive experience using Agile and DevOps methodologies to build product infrastructure and shared services at scale.
  • •Experience running large-scale infrastructure platforms for a SaaS/Cloud service in a public cloud, preferably AWS (multi-cloud experience is a plus).
  • •Strong expertise in cloud-native architectures, Kubernetes, Terraform (IaC), and CI/CD pipelines.
  • •Deep experience building and operating observability platforms and monitoring tools (Grafana, Splunk, APM) at scale.
Experience:3+ yearsSaaSCloudDevOpsMicroservices
Skills:LeadershipPeople managementVerbal communicationWritten communicationInterpersonal skills
Tech Stack:AWSAvailability zonesRegionsKubernetesK8sCI/CDTerraformCloud-native architecturesObservabilityGrafanaSplunkAPMInfrastructure as codeMicroservicesAgileDevOpsSDLCPaaSAutomationEdge networking

Eligibility

Nationality:US National

Company Brief

Okta
Provides identity and access management cloud solutions that help organizations secure and manage user authentication, single sign-on, multi-factor authentication, and lifecycle management across applications and devices.
Industry: Cybersecurity
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: San Francisco, United States
Founded: 2009
WebsiteLinkedIn