Site Reliability Engineer

Darktrace
Cambridge
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureSkills: ["Go","Python","AWS","GCP","Azure","Kubernetes","Observability","Monitoring","Distributed tracing"]

Join Darktrace as a Site Reliability Engineer focused on a core reliability domain, driving standards and tooling across SRE, Platform Engineering, and DevSecOps. You’ll own cross-cutting reliability challenges, shape runbooks and best practices, and partner with security and dev teams to improve observability, performance, and resilience at scale.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Darktrace
Darktrace
2 months ago

Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 7 days agoStatus: Live

Job Summary

Join Darktrace as a Site Reliability Engineer focused on a core reliability domain, driving standards and tooling across SRE, Platform Engineering, and DevSecOps. You’ll own cross-cutting reliability challenges, shape runbooks and best practices, and partner with security and dev teams to improve observability, performance, and resilience at scale.
Location: Cambridge
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure

Key Responsibilities

  • •Act as the subject matter expert in your chosen reliability domain
  • •Define and implement standards, frameworks, and best practices across SRE, Platform Engineering, and DevSecOps
  • •Stay current with industry trends and bring innovative ideas into the organisation
  • •Design and implement solutions to complex, cross-cutting reliability challenges
  • •Build tooling, automation, and frameworks to improve system resilience and scalability
  • •Lead deep-dive investigations into systemic issues and drive long-term fixes
  • •Partner with Platform Engineering to embed your domain within the internal developer platform
  • •Collaborate with DevSecOps to integrate security, compliance, and resilience practices
  • •Contribute to cross-team initiatives that improve reliability across the stack
  • •Play a key role in incident response and on-call rotations, and develop runbooks and training materials

Pay and Benefits

Perks:Health InsuranceLife InsurancePensionHolidayBirthday OffCycle To

Key Requirements

  • •Proven experience in Site Reliability Engineering, DevOps, or infrastructure engineering
  • •Deep expertise in at least one domain: observability/monitoring, performance engineering, data infrastructure reliability, security-focused SRE, or network reliability
  • •Strong programming skills (Go, Python, or similar)
  • •Experience with cloud platforms (AWS, GCP, Azure) and Kubernetes
  • •Strong communication skills and ability to identify and prioritise high-impact work independently
Experience:CloudSREDevOpsCybersecurity
Skills:GoPythonAWSGCPAzureKubernetesObservabilityMonitoringDistributed tracing
Tech Stack:GoPythonAWSGCPAzureKubernetes

Company Brief

Darktrace
Develops AI-driven cybersecurity solutions that detect, respond to, and neutralize cyber threats across cloud, email, network, and endpoint environments using machine learning and anomaly detection to protect enterprises globally.
Industry: Cybersecurity
Company Size: Enterprise (1,001+ employees)
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Cambridge, United Kingdom
Founded: 2013
WebsiteLinkedIn