Senior Site Reliability Engineer
Cambridge
Workplace: RemoteFull timeUSD 121,400 - 218,600 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 5+ yearsEducation: bachelorsSkills: ["Collaboration","Ownership","Problem-solving","Cross-functional coordination","Incident management"]Build and scale SRE tooling for Akamai’s next-generation dedicated AI hardware infrastructure. Partner with product teams to ensure reliability, scalability, and performance across globe-spanning systems by defining and defending key performance indicators. Develop infrastructure-as-code utilities in Python, automate provisioning and break-fix workflows, implement telemetry and Prometheus/Grafana dashboards with anomaly detection, and lead 24x7 incident response using PagerDuty and Slack while coordinating vendors and field technicians.
Loading
Loading job details...
Preparing the role view and application actions.

