Senior II Site Reliability Engineer

Akamai Technologies
Krakow
Workplace: RemoteFull timeFunction: DevOps, Cloud & InfrastructureEducation: bachelorsSkills: ["Collaboration","Mentorship","Troubleshooting","Automation","Reliability focus"]

Ensure the operation and uptime of Compute services and infrastructure by supervising critical systems and partnering with cross-functional teams to build tooling that monitors and improves reliability. Improve Akamai’s Compute Cloud Interface platform for faster error detection and remediation, develop automation to reduce toil, and participate in on-call rotations. Support and mentor fellow SREs while deploying and maintaining internal platform tools.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Akamai Technologies
Akamai Technologies
1 day ago

Senior II Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 19 hours agoStatus: Live
Reposted: similar role first listed 4 months ago

Job Summary

Ensure the operation and uptime of Compute services and infrastructure by supervising critical systems and partnering with cross-functional teams to build tooling that monitors and improves reliability. Improve Akamai’s Compute Cloud Interface platform for faster error detection and remediation, develop automation to reduce toil, and participate in on-call rotations. Support and mentor fellow SREs while deploying and maintaining internal platform tools.
Location: Krakow
Workplace: Remote
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Provide support and mentorship for other SRE engineers within the team
  • •Deploy and maintain the platform and internal tools
  • •Improve the Compute Cloud Interface platform to speed error detection and remediation, enhancing performance and reliability
  • •Develop and improve automation to support daily activities and reduce toil
  • •Participate in on-call rotations and collaborate on troubleshooting and resolving escalations and incidents

Key Requirements

  • •Extensive relevant experience and a Bachelor's degree in Computer Science or equivalent
  • •Experience automating with Python and/or Golang, plus scripting with bash
  • •Knowledge of systems reliability best practices including observability and monitoring with SLO adherence
  • •Experience with configuration management tools such as SaltStack, Terraform, and Ansible, and CI/CD solutions such as Jenkins
  • •Hands-on Linux administration and container platforms like Docker, plus monitoring/logging tools (Prometheus, Grafana, Loki) and tools such as nginx/envoy/haproxy and Redis
Experience:Site reliability engineeringSRECloudObservabilityAutomation
Education:Bachelor's
Skills:CollaborationMentorshipTroubleshootingAutomationReliability focus
Languages:US
Tech Stack:PythonGolangBashSaltStackTerraformAnsibleJenkinsLinuxDockerPrometheusGrafanaLokiNginxEnvoyHaproxyRedis

Company Brief

Akamai Technologies
Provides a global content delivery network (CDN) and cloud services to improve web and application performance, security, and delivery for enterprises, media companies, and cloud providers.
Industry: Cloud Computing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Cambridge, United States
Founded: 1998
WebsiteLinkedIn