Site Reliability Engineer – NS London

BAE Systems
London
Workplace: HybridFull timeFunction: DevOps, Cloud & InfrastructureSkills: ["Problem-solving","Innovation","Continuous improvement","Troubleshooting","Collaboration","Learning mindset"]

Build and support essential mission-critical systems as part of an SRE team, replacing manual operations with automation to improve availability, performance, and stability. Join a 24/7 on-call rota to resolve production incidents, design monitoring and bespoke observability tools, and partner with development teams on system design best practices. The role also includes contributing to the wider DevOps/SRE community and adopting new technologies to strengthen scalability and resilience.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
BAE Systems
BAE Systems
4 months ago

Site Reliability Engineer – NS London

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 12 hours agoStatus: Live

Job Summary

Build and support essential mission-critical systems as part of an SRE team, replacing manual operations with automation to improve availability, performance, and stability. Join a 24/7 on-call rota to resolve production incidents, design monitoring and bespoke observability tools, and partner with development teams on system design best practices. The role also includes contributing to the wider DevOps/SRE community and adopting new technologies to strengthen scalability and resilience.
Location: London
Workplace: Hybrid
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure

Key Responsibilities

  • •Support and maintain essential services for core mission applications, proactively improving availability, performance, and stability.
  • •Join a 24/7 on-call rota to support critical production systems out of business hours and resolve incidents.
  • •Automate repetitive work and find innovative solutions to improve reliability, aiming to reduce manual operations time.
  • •Design and deploy monitoring products and build bespoke observability tools to provide intelligent insights and demonstrate improvements.
  • •Advise development teams on good practices for designing and building scalable, resilient systems and participate in the internal DevOps/SRE community.

Pay and Benefits

Perks:Hybrid WorkOvertime

Key Requirements

  • •Experience supporting and maintaining production services with a focus on availability, performance, and stability.
  • •Strong software development skills in web technologies and object-oriented programming.
  • •Working knowledge of databases including Oracle SQL, Mongo, or PostgreSQL.
  • •Comfort with Linux/Windows command lines (e.g., Bash and PowerShell) and monitoring tools such as Grafana, Prometheus, ELK, or Splunk.
  • •Understanding microservices architecture and container platforms such as Docker and Kubernetes (including OpenShift).
Experience:DevOpsSREAgileMonitoringMicroservicesContainersObservability
Skills:Problem-solvingInnovationContinuous improvementTroubleshootingCollaborationLearning mindset
Tech Stack:LinuxWindowsBashPowerShellOracle SQLMongoPostgreSQLGrafanaPrometheusELKSplunkAgileAtlassianITILMicroservicesDockerOpenShiftKubernetes

Eligibility

Security Clearance:EDV

Company Brief

BAE Systems
Multinational defence, security and aerospace company designing, manufacturing and supporting military and civilian systems including naval ships, submarines, aircraft, cyber solutions and electronic systems for governments and commercial customers.
Industry: Defense Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: London, United Kingdom
Founded: 1999
WebsiteLinkedIn