Site Reliability Engineer, Infrastructure Platforms — UK (Intermediate to Senior Staff)

GitLab
United Kingdom
Workplace: RemoteFull timeFunction: DevOps, Cloud & InfrastructureSkills: ["Growth mindset","Strong written communication","Ability to operate as a manager-of-one","Ownership","Structured troubleshooting"]

Build and operate GitLab’s production infrastructure to keep user-facing services reliable, scalable, and efficient. You’ll create automation and infrastructure-as-code workflows, run and troubleshoot systems on Kubernetes, and improve CI/CD and GitOps delivery. Own incident response and post-incident learnings, advance observability with metrics/logs/SLOs, and maintain runbooks and architecture decisions across Infrastructure Platforms teams.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
GitLab
GitLab
1 day ago

Site Reliability Engineer, Infrastructure Platforms — UK (Intermediate to Senior Staff)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 15 hours agoStatus: Live

Job Summary

Build and operate GitLab’s production infrastructure to keep user-facing services reliable, scalable, and efficient. You’ll create automation and infrastructure-as-code workflows, run and troubleshoot systems on Kubernetes, and improve CI/CD and GitOps delivery. Own incident response and post-incident learnings, advance observability with metrics/logs/SLOs, and maintain runbooks and architecture decisions across Infrastructure Platforms teams.
Location: United Kingdom
Workplace: Remote
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Keep user-facing services and production systems reliable, scalable, and efficient.
  • •Build automation and tooling that reduces toil using infrastructure-as-code-driven workflows.
  • •Operate and troubleshoot production systems on Kubernetes, including deployments, rollouts, and scaling.
  • •Write and maintain infrastructure as code, shipping changes safely through CI/CD and GitOps.
  • •Participate in on-call, triage alerts, perform incident response and post-incident reviews, and improve runbooks, observability, and escalation practices.

Pay and Benefits

Perks:Paid LeaveEquityParental LeaveLearning Budget

Key Requirements

  • •Experience keeping production systems reliable with strong operations mindset and real software engineering practice.
  • •Build net-new infrastructure tooling and automation (e.g., Terraform modules, Kubernetes operators/controllers, or automation/services written from scratch).
  • •Ability to read, debug, and reason about code; experience with Go and/or Ruby.
  • •Hands-on infrastructure as code and Kubernetes ecosystem knowledge at a depth appropriate to your level.
  • •Hands-on experience with at least one major cloud provider (GCP or AWS), plus observability practices (metrics, logging, alerting, SLOs/SLIs) and comfort in on-call/incident response.
Skills:Growth mindsetStrong written communicationAbility to operate as a manager-of-oneOwnershipStructured troubleshooting
Tech Stack:KubernetesTerraformGoRubyAWSGCPCI/CDGitOpsInfrastructure as codeObservabilityMetricsLogsAlertingSLOsSLIs

Company Brief

GitLab
Provides a single application for the complete DevSecOps lifecycle, offering source code management, CI/CD, security, and collaboration tools to help teams deliver software faster and more securely.
Industry: Developer Tools
Company Size: Enterprise (1,001+ employees)
Revenue: USD 250M to 500M
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: San Francisco, United States
Founded: 2011
WebsiteLinkedIn