Principal Site Reliability Engineer Lead

Akamai Technologies
Krakow
Workplace: RemoteFull timeFunction: DevOps, Cloud & InfrastructureExperience: 10+ yearsEducation: bachelorsSkills: ["Methodical problem-solving","Technical leadership","Mentoring","Collaboration","On-call incident response"]

Provide technical and architectural leadership to design, build, and scale a Next Generation Control Plane for Akamai Cloud. Lead the transformation toward a highly available, resilient, distributed microservices architecture running on Kubernetes and other cloud-native technologies. Own platform reliability through SLOs and KPIs, collaborate across Engineering, Product, and Support, and participate in on-call rotations guiding restoration and repair of service-impacting issues.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Akamai Technologies
Akamai Technologies
3 days ago

Principal Site Reliability Engineer Lead

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 27 minutes agoStatus: Live

Job Summary

Provide technical and architectural leadership to design, build, and scale a Next Generation Control Plane for Akamai Cloud. Lead the transformation toward a highly available, resilient, distributed microservices architecture running on Kubernetes and other cloud-native technologies. Own platform reliability through SLOs and KPIs, collaborate across Engineering, Product, and Support, and participate in on-call rotations guiding restoration and repair of service-impacting issues.
Location: Krakow
Workplace: Remote
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure

Key Responsibilities

  • •Contribute to the design, development, and operation of a Golden Path platform built on Kubernetes and other cloud-native technologies.
  • •Provide leadership, support, and mentoring to your team.
  • •Create and maintain SLOs and KPIs, collaborating with Engineering, Product, and Support teams.
  • •Participate in on-call rotations and guide restoration and repair of service-impacting issues.
  • •Lead modernization efforts by designing, building, and scaling a highly available, resilient, distributed microservices architecture.

Key Requirements

  • •10+ years of relevant experience and a Bachelor's degree in Computer Science (or similar) or equivalent experience.
  • •Deep experience building and operating highly available, fault-tolerant, scalable production services using Kubernetes and other cloud-native technologies.
  • •Strong understanding of Linux internals, especially containerization and networking.
  • •Code-first operations to eliminate toil via code fixes and automation.
  • •Experience with observability tooling (e.g., OpenTelemetry, Prometheus, Grafana, Loki) and infrastructure-as-code tools (e.g., Crossplane, Pulumi, Terraform, Ansible).
Experience:10+ yearsCloud-nativeKubernetesMicroservicesInfrastructure as codeObservability
Education:Bachelor's in Computer Science
Skills:Methodical problem-solvingTechnical leadershipMentoringCollaborationOn-call incident response
Tech Stack:KubernetesGoPythonBashLinuxContainerizationNetworkingCrossplanePulumiTerraformAnsibleOpenTelemetryPrometheusGrafanaLokiMicroservicesObservability

Company Brief

Akamai Technologies
Provides a global content delivery network (CDN) and cloud services to improve web and application performance, security, and delivery for enterprises, media companies, and cloud providers.
Industry: Cloud Computing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Cambridge, United States
Founded: 1998
WebsiteLinkedIn