Site Reliability Engineer

Trainline
London
Workplace: HybridFull timeGBP 55,000 - 63,000 annuallyFunction: DevOps, Cloud & InfrastructureSkills: ["Communication","Collaboration","Problem-solving","Rapid learning","Stakeholder management"]

Mid-level Site Reliability Engineer at Trainline, focusing on reliability, observability, and incident response. You’ll work with AWS-based cloud infrastructure, CI/CD, and DevOps practices to improve monitoring, post-incident reviews, and on-call resilience, collaborating with product engineering to ensure safe deployments and excellent service reliability.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Trainline
Trainline
6 months ago

Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 13 hours agoStatus: Live

Job Summary

Mid-level Site Reliability Engineer at Trainline, focusing on reliability, observability, and incident response. You’ll work with AWS-based cloud infrastructure, CI/CD, and DevOps practices to improve monitoring, post-incident reviews, and on-call resilience, collaborating with product engineering to ensure safe deployments and excellent service reliability.
Location: London
Workplace: Hybrid
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Develop, maintain, and improve observability across the Trainline platform using metrics, logs, events, and traces to support detection and diagnosis.
  • •Participate in production incident response, investigate, communicate, and coordinate service restoration.
  • •Contribute to post-incident reviews and follow-up actions to enhance reliability, scalability, and resilience.
  • •Take part in the SRE on-call rotation and ensure services are operationally ready for safe deployments.
  • •Design and implement tooling and pipelines to strengthen monitoring, alerting, and data surfaced during live incidents.

Pay and Benefits

Salary: GBP 55,000 - 63,000 annually
Perks:Health InsurancePension

Key Requirements

  • •Strong production experience in site reliability engineering or SRE/DevOps roles.
  • •Hands-on experience with observability tooling (e.g., New Relic, ELK, Grafana) and cloud providers (preferably AWS).
  • •Experience scripting in at least one language (preferably Python) and with infrastructure-as-code tooling (Terraform, GitHub Actions).
  • •Understanding of SRE concepts such as SLI, SLO, and error budgets, and experience with incident response and post-incident reviews.
  • •Ability to design and implement monitoring, alerting, and time-series data to support rapid detection and diagnosis.
Skills:CommunicationCollaborationProblem-solvingRapid learningStakeholder management
Languages:English
Tech Stack:AWSNew RelicELKGrafanaIncident.ioDockerECSTerraformGitHub ActionsPython

Company Brief

Trainline
Digital rail and coach ticketing platform offering online booking, mobile tickets, journey planning, and real-time travel information across the UK and Europe. Aggregates fares from multiple operators and provides retail and distribution services for rail travel.
Industry: TravelTech
Company Size: Enterprise (1,001+ employees)
Revenue: USD 100M to 250M
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: London, United Kingdom
Founded: 1997
WebsiteLinkedIn