Senior Site Reliability Engineer - Platform Reliability (Resilience)

Elastic
Greece
Workplace: OnsiteFull timeEUR 71,200 - 92,400 annuallyFunction: DevOps, Cloud & InfrastructureSkills: ["Collaboration","Operational excellence","Customer-first mindset","Problem-solving","Mentoring"]

Join the Platform Engineering SRE team designing and operating the multi-cloud platform that hosts Elastic Cloud Hosted and Serverless. You’ll lead reliability initiatives for automated system engineering, grow global platform infrastructure to meet scaling demands, and manage major-incident response using a follow-the-sun on-call rotation. The role emphasizes customer-first operational problem solving and building tooling and automation to prevent repeated impact.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Elastic
Elastic
1 day ago

Senior Site Reliability Engineer - Platform Reliability (Resilience)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 2 hours agoStatus: Live

Job Summary

Join the Platform Engineering SRE team designing and operating the multi-cloud platform that hosts Elastic Cloud Hosted and Serverless. You’ll lead reliability initiatives for automated system engineering, grow global platform infrastructure to meet scaling demands, and manage major-incident response using a follow-the-sun on-call rotation. The role emphasizes customer-first operational problem solving and building tooling and automation to prevent repeated impact.
Location: Greece
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Lead technical initiatives to automate system engineering efforts that guarantee global infrastructure reliability.
  • •Grow and maintain global platform infrastructure by developing software, tooling, and automations for scaling.
  • •Respond to and prevent repeated customer impact through major incident response and prioritized problem management.
  • •Participate in a follow-the-sun on-call rotation for operational coverage during (mostly) local working hours.
  • •Champion a collaborative, inclusive environment focused on operational excellence and uplifting others.

Pay and Benefits

Salary: EUR 71,200 - 92,400 annually
Perks:Health InsuranceParental LeaveVolunteer TimePaid Leave

Key Requirements

  • •Demonstrated experience driving platform reliability with a customer-first, operational problem-solving mindset.
  • •Background in software engineering to identify, implement, and deliver solutions with engineering partners.
  • •Experience with public cloud and managed Kubernetes services.
  • •Experience leading and improving alerting and major incident management processes, metrics, and systems to diagnose and quantify impact.
  • •Strong system administration skills with Linux on distributed systems at scale.
Experience:SaaSPublic cloudInfrastructure-as-codeKubernetesDistributed systems
Skills:CollaborationOperational excellenceCustomer-first mindsetProblem-solvingMentoring
Languages:English
Tech Stack:GolangCrossplaneTerraformKubernetesDockerLinuxElastic StackGraphitePrometheusInflux

Company Brief

Elastic
Builds the Elastic Stack (Elasticsearch, Kibana, Beats, Logstash) and provides search, observability, and security solutions that enable organizations to search, analyze, and protect data in real time across applications, infrastructure, and enterprises.
Industry: Data Infrastructure
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Amsterdam, Netherlands
Founded: 2012
WebsiteLinkedIn