Senior Site Reliability Engineer (Hosted Infra)

Elastic
Australia
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureSkills: ["Systems thinking","Root cause analysis","Documentation","Communication","Code review"]

Build and automate the infrastructure that powers Elastic Cloud across multiple cloud providers and thousands of regions. You’ll improve host reliability and lifecycle, strengthen observability with alerting and monitoring for incident prevention, and scale global systems using Infrastructure as Code and purpose-built tooling. Collaborate on software engineering for internal services, contribute to code reviews, and participate in a balanced on-call rotation with runbooks and postmortems.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Elastic
Elastic
1 day ago

Senior Site Reliability Engineer (Hosted Infra)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 19 hours agoStatus: Live

Job Summary

Build and automate the infrastructure that powers Elastic Cloud across multiple cloud providers and thousands of regions. You’ll improve host reliability and lifecycle, strengthen observability with alerting and monitoring for incident prevention, and scale global systems using Infrastructure as Code and purpose-built tooling. Collaborate on software engineering for internal services, contribute to code reviews, and participate in a balanced on-call rotation with runbooks and postmortems.
Location: Australia
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Engineering software to automate large-scale systems by building internal tools and services.
  • •Optimizing reliability and the lifecycle of hosts across multiple cloud providers.
  • •Strengthening observability by crafting alerting and monitoring to prevent incidents.
  • •Scaling global infrastructure and evolving infrastructure management processes to meet demand.
  • •Participating in on-call rotation, improving runbooks, and contributing to postmortems and reliability improvements.

Pay and Benefits

Perks:Health InsurancePaid LeaveParental LeaveDonation MatchVolunteer Time

Key Requirements

  • •Production experience operating large-scale cloud compute (hundreds of hosts or more) via automated workflows.
  • •Deep experience with Linux systems and debugging at the OS level.
  • •Experience building software with Golang and providing constructive code review feedback.
  • •Proficiency running containerized workloads in production.
  • •Customer-first, systems-thinking approach focused on root causes, with clear documentation and communication.
Skills:Systems thinkingRoot cause analysisDocumentationCommunicationCode review
Languages:English
Tech Stack:GolangLinuxInfrastructure as Code (IaC)ContainersTerraformOpenTofuPuppetOpenVoxAnsibleArgo CDArgo WorkflowsCUEDockerKubernetesUbuntuUbuntu Live PatchElastic StackGraphitePrometheusInflux

Company Brief

Elastic
Builds the Elastic Stack (Elasticsearch, Kibana, Beats, Logstash) and provides search, observability, and security solutions that enable organizations to search, analyze, and protect data in real time across applications, infrastructure, and enterprises.
Industry: Data Infrastructure
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Amsterdam, Netherlands
Founded: 2012
WebsiteLinkedIn