Senior Site Reliability Engineer
Akamai Technologies
Krakow
Workplace: RemoteFull timeFunction: DevOps, Cloud & InfrastructureSkills: ["Collaboration","Analytical thinking","Ownership","Incident response","Problem-solving"]Ensure reliability and operational readiness for AI hardware and software systems across regional data centers. Partner with product teams to improve scalability, performance, and uptime by defining KPIs, monitoring breaches, and driving incident response. Build and scale Python infrastructure-as-code tooling, automate workflows across JIRA/Siebel/PagerDuty, and design observability pipelines with Prometheus/Grafana and OpenTelemetry/Loki. Join 24x7x365 on-call and coordinate with vendors and field technicians to maintain service availability.

