Lead Site Reliability Engineer

Glean
Palo Alto
Workplace: HybridFull timeUSD 200,000 - 260,000 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 8+ yearsEducation: bachelorsSkills: ["Leadership","Mentorship","Problem-solving","On-call","Communication"]

Lead and mentor a team of Site Reliability Engineers to ensure high availability and reliability of cloud-based services. Drive technical strategy, incident management, automation, and monitoring, while optimizing performance and security across hybrid cloud environments.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Glean
Glean
7 months ago

Lead Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 1 hour agoStatus: Live

Job Summary

Lead and mentor a team of Site Reliability Engineers to ensure high availability and reliability of cloud-based services. Drive technical strategy, incident management, automation, and monitoring, while optimizing performance and security across hybrid cloud environments.
Location: Palo Alto
Workplace: Hybrid
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure

Key Responsibilities

  • •Technical Leadership and Mentorship: Drive technical excellence, establish best practices for incident management, performance optimization, and automation, and influence cross-team collaborations.
  • •Ensure High Availability: Build and maintain resilient cloud architectures, monitor performance, and proactively resolve bottlenecks or failure points.
  • •Incident Management: Participate in primary on-call rotation; promote a blameless postmortem culture and optimize on-call processes.
  • •Automation and Tooling: Develop and maintain automation scripts, tools, and processes to streamline deployment, monitoring, and management tasks for scalable cloud operations.
  • •Monitoring and Alerting: Design advanced monitoring, create dashboards, and set up alerts to proactively detect and respond to issues.

Pay and Benefits

Salary: USD 200,000 - 260,000 annually
Perks:MedicalVisionDental401kHome OfficeEducation StipendWellness Stipend

Key Requirements

  • •8+ years of experience in a senior-level role within Site Reliability Engineering or similar role, particularly in managing cloud-based services and infrastructure.
  • •5+ years of experience with software development in one or more programming languages.
  • •3+ years of experience managing people or teams, leading projects, and designing, analyzing, and troubleshooting distributed systems running in Cloud.
  • •Strong knowledge of cloud platforms such as Google Cloud Platform, AWS, or Azure.
  • •Practical experience with containerization technologies, including Docker and Kubernetes. Familiarity with infrastructure as code tools like Terraform is essential.
Experience:8+ yearsCloud computingSREDistributed systems
Education:Bachelor's
Skills:LeadershipMentorshipProblem-solvingOn-callCommunication
Languages:English
Tech Stack:Google Cloud PlatformAWSAzureDockerKubernetesTerraformMonitoringOn-call tooling

Company Brief

Glean
Provides an enterprise search and knowledge discovery platform that helps organizations find information across apps, drives, and internal tools, improving employee productivity and onboarding through AI-powered search and relevance ranking.
Industry: Enterprise Software
Company Size: Large (251 to 1,000 employees)
Growth: Scaleup
Headquarters: San Francisco, United States
Founded: 2017
WebsiteLinkedIn