Staff Site Reliability Engineer

Harvey
Bengaluru
Workplace: HybridFull timeFunction: Software EngineeringExperience: 10+ yearsSkills: ["Leadership","Mentorship","Problem-solving","Attention to detail","Ownership"]

Staff SRE geared role focused on ensuring reliability, scalability, and performance of Harvey's AI-enabled platform. You will own monitoring, incident response, automation of operational tasks, and security/reliability best practices across 50+ regions, while mentoring teams and optimizing infrastructure costs. This position sits at the intersection of infrastructure and product, shaping a resilient, high-uptime system as Harvey scales globally.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Harvey
Harvey
4 months ago

Staff Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 50 minutes agoStatus: Live

Job Summary

Staff SRE geared role focused on ensuring reliability, scalability, and performance of Harvey's AI-enabled platform. You will own monitoring, incident response, automation of operational tasks, and security/reliability best practices across 50+ regions, while mentoring teams and optimizing infrastructure costs. This position sits at the intersection of infrastructure and product, shaping a resilient, high-uptime system as Harvey scales globally.
Location: Bengaluru
Workplace: Hybrid
Employment Type: Full time
Job Function: Software Engineering

Key Responsibilities

  • •Design, implement, and manage monitoring, alerting, and infrastructure resources (compute, storage, networking) across 50+ global regions
  • •Lead incident management processes, including postmortems, root cause analyses, and driving actionable improvements
  • •Automate operational tasks and workflows, building tools and processes for capacity planning, graceful rollouts, and safe data access to maintain high reliability and reduce manual intervention
  • •Establish best practices for security, compliance, and reliability and collaborate across teams to drive these principles throughout the software lifecycle
  • •Optimize infrastructure costs through strategic capacity planning and build-versus-buy decisions while maintaining system performance, reliability, and functionality

Key Requirements

  • •10+ years of experience in Site Reliability Engineering or similar roles supporting production environments, with proven ability to mentor and guide technical teams
  • •Expertise in infrastructure as code(IaC) tools (Pulumi, Terraform, CloudFormation, etc.)
  • •Deep familiarity with observability tools (Datadog, Sentry, etc.) and incident response practices (PagerDuty, IncidentIO, etc.)
  • •Proficiency with cloud infrastructure platforms (Azure, GCP, AWS, etc.)
  • •Strong programming skills (Python, Bash, Go, or similar languages)
Experience:10+ yearsSite Reliability EngineeringSREInfrastructure as codeCloud infrastructure
Skills:LeadershipMentorshipProblem-solvingAttention to detailOwnership
Tech Stack:PulumiTerraformCloudFormationDatadogSentryPagerDutyIncidentIOAzureGCPAWSPythonBashGoKubernetesCI/CDDockerNetworkingDatabasesCloud security

Eligibility

Work Authorization:Authorization required. Sponsorship not provided.

Company Brief

Harvey
Harvey builds domain-specific generative AI for legal and professional services, automating contract analysis, due diligence, compliance, and litigation workflows for law firms and corporate legal teams.
Industry: LegalTech
Company Size: Large (251 to 1,000 employees)
Revenue: USD 100M to 250M
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2022
Glassdoor
Glassdoor: 4.1
WebsiteLinkedInGlassdoor