Senior Site Reliability Engineer

Okta
Dublin
Workplace: HybridFull timeEUR 76,000 - 104,500 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 5+ yearsEducation: bachelorsSkills: ["Communication","Collaboration","Mentoring"]

Build and operate a Kubernetes-based internal platform that hosts business workflows, evolving SRE into a Platform as a Service (PaaS). You’ll develop automation and “golden paths,” ensure production readiness for AI-driven workflows, and lead incident ownership through troubleshooting, postmortems, and reliability practices (SLIs/SLOs, runbooks). Improve CI/CD and infrastructure-as-code, strengthen observability, and mentor engineers across the Dublin/EMEA team and global SRE peers.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Okta
Okta
19 hours ago

Senior Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 14 hours agoStatus: Live

Job Summary

Build and operate a Kubernetes-based internal platform that hosts business workflows, evolving SRE into a Platform as a Service (PaaS). You’ll develop automation and “golden paths,” ensure production readiness for AI-driven workflows, and lead incident ownership through troubleshooting, postmortems, and reliability practices (SLIs/SLOs, runbooks). Improve CI/CD and infrastructure-as-code, strengthen observability, and mentor engineers across the Dublin/EMEA team and global SRE peers.
Location: Dublin
Workplace: Hybrid
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Build, operate, and improve a Kubernetes platform (cluster management, networking, security posture, and day-2 operations).
  • •Develop self-service capabilities, golden paths, and automation to enable workflow onboarding and operation with less platform expertise.
  • •Ensure production readiness for AI-driven internal workflows at scale, including incident troubleshooting and postmortems.
  • •Uphold reliability practices such as SLIs/SLOs, error budgets, and operational runbooks.
  • •Improve CI/CD pipelines and infrastructure-as-code practices, and enhance observability and monitoring coverage.

Pay and Benefits

Salary: EUR 76,000 - 104,500 annually
Perks:Health InsurancePaid LeaveParental Leave

Key Requirements

  • •5+ years of experience in Site Reliability Engineering, Platform Engineering, or Infrastructure Engineering.
  • •Hands-on Kubernetes experience in production, including deployment, networking, security, and troubleshooting.
  • •Operating large-scale internal infrastructure platforms in a public cloud, preferably AWS.
  • •Infrastructure-as-code (Terraform) and CI/CD pipeline experience, including change/release management.
  • •Experience with observability and monitoring tools such as Grafana and Splunk (or equivalent), plus a CS degree or related experience.
Experience:5+ yearsPlatform engineeringInfrastructure engineeringAWSKubernetesAI/ML
Education:Bachelor's
Skills:CommunicationCollaborationMentoring
Certifications:CKACKSCKAD
Tech Stack:AWSKubernetesTerraformCI/CDGrafanaSplunkAPMIstioLinkerdArgoCDFluxGitOpsGitHubGitLabVector databasesModel routingInference infrastructure

Company Brief

Okta
Provides identity and access management cloud solutions that help organizations secure and manage user authentication, single sign-on, multi-factor authentication, and lifecycle management across applications and devices.
Industry: Cybersecurity
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: San Francisco, United States
Founded: 2009
WebsiteLinkedIn