Senior Site Reliability Engineer (Observability & Analytics) – Platform Infra

Elastic
Canada
Workplace: RemoteFull timeCAD 138,300 - 185,900 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 5+ yearsSkills: ["Incident response","Mentoring","Risk awareness","Written communication","Verbal communication"]

Own end-to-end delivery for observability and analytics platform infrastructure, supporting 200+ hosted deployments across cloud regions. Strengthen shared Elastic Cloud infrastructure using Infrastructure as Code with Terraform, Python, and Go, and provide 24/7 incident response with RCA-driven fixes. Review critical production changes, mentor engineers, and improve runbooks and operational processes to reduce on-call load, while maintaining a security-conscious approach.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Elastic
Elastic
1 day ago

Senior Site Reliability Engineer (Observability & Analytics) – Platform Infra

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 6 hours agoStatus: Live

Job Summary

Own end-to-end delivery for observability and analytics platform infrastructure, supporting 200+ hosted deployments across cloud regions. Strengthen shared Elastic Cloud infrastructure using Infrastructure as Code with Terraform, Python, and Go, and provide 24/7 incident response with RCA-driven fixes. Review critical production changes, mentor engineers, and improve runbooks and operational processes to reduce on-call load, while maintaining a security-conscious approach.
Location: Canada
Workplace: Remote
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Own end-to-end delivery of moderate-to-high complexity projects on the team roadmap with minimal day-to-day direction.
  • •Operate and harden shared Elastic Cloud infrastructure (ECH, ECE, ECK) as Infrastructure as Code, including writing and reviewing Terraform, Python, and Go changes.
  • •Participate in a 24/7 on-call rotation by responding to incidents, driving resolution, and writing RCAs/postmortems for lasting fixes.
  • •Review others’ code and designs and serve as a trusted second set of eyes on production changes to critical infrastructure.
  • •Mentor less experienced engineers and improve runbooks, documentation, and operational processes to reduce on-call load.

Pay and Benefits

Salary: CAD 138,300 - 185,900 annually
Equity and Bonus:Equity
Perks:Health InsurancePaid LeaveParental LeaveRrspVolunteer Time

Key Requirements

  • •5+ years of SRE, platform engineering, or infrastructure engineering experience.
  • •Proficiency with Terraform, including owning large multi-workspace configurations in a team setting.
  • •Strong software engineering fundamentals in Python; comfort with Go is a plus.
  • •Deep Linux systems knowledge and experience operating containerized workloads in production.
  • •Experience on a 24/7 on-call rotation, resolving incidents under pressure, and writing RCAs that hold up under review.
Experience:5+ yearsInfrastructure engineeringPlatform engineeringSRE
Skills:Incident responseMentoringRisk awarenessWritten communicationVerbal communication
Languages:English
Tech Stack:TerraformPythonGoLinuxContainersECHECEECKKubernetesElastic StackElasticsearchLogstashBeatsKibanaArgoCDHelmKyvernoVaultTeleportPuppet

Company Brief

Elastic
Builds the Elastic Stack (Elasticsearch, Kibana, Beats, Logstash) and provides search, observability, and security solutions that enable organizations to search, analyze, and protect data in real time across applications, infrastructure, and enterprises.
Industry: Data Infrastructure
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Amsterdam, Netherlands
Founded: 2012
WebsiteLinkedIn