Site Reliability Engineer, Intermediate to Senior Staff — Infrastructure Platforms

GitLab
United States
Workplace: RemoteFull timeUSD 126,400 - 314,400 annuallyFunction: DevOps, Cloud & InfrastructureSkills: ["Growth mindset","Strong written communication","Ability to operate independently","Incident response troubleshooting","Automation mindset"]

Build and operate reliable, scalable production systems for GitLab.com. Create automation and infrastructure-as-code tooling to reduce toil, and run Kubernetes-based services across deployments, rollouts, and scaling. Improve CI/CD and GitOps workflows, participate in on-call and incident response, and advance observability using metrics, logs, and SLOs. Document runbooks and architecture decisions, turning post-incident learnings into repeatable improvements.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
GitLab
GitLab
1 month ago

Site Reliability Engineer, Intermediate to Senior Staff — Infrastructure Platforms

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 5 hours agoStatus: Live

Job Summary

Build and operate reliable, scalable production systems for GitLab.com. Create automation and infrastructure-as-code tooling to reduce toil, and run Kubernetes-based services across deployments, rollouts, and scaling. Improve CI/CD and GitOps workflows, participate in on-call and incident response, and advance observability using metrics, logs, and SLOs. Document runbooks and architecture decisions, turning post-incident learnings into repeatable improvements.
Location: United States
Workplace: Remote
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure

Key Responsibilities

  • •Keep user-facing services and production systems reliable, scalable, and efficient.
  • •Build automation and tooling using infrastructure-as-code-driven workflows to reduce toil.
  • •Operate and troubleshoot production systems on Kubernetes (deployments, rollouts, scaling).
  • •Write and maintain infrastructure as code and ship changes safely via CI/CD and GitOps.
  • •Participate in on-call, triage alerts, improve runbooks, and support incident response and post-incident reviews.

Pay and Benefits

Salary: USD 126,400 - 314,400 annually
Equity and Bonus:Equity
Perks:Paid LeaveEquityLearning BudgetParental Leave

Key Requirements

  • •Experience keeping production systems reliable, combining operations mindset with software engineering practice.
  • •Build net-new infrastructure tooling and automation (e.g., Terraform modules, Kubernetes operators/controllers, or automation/services written from scratch).
  • •Ability to read, debug, and reason about code; experience with Go and/or Ruby is a plus.
  • •Experience with infrastructure as code and Kubernetes ecosystem at a depth appropriate to your level.
  • •Hands-on experience with at least one major cloud provider (GCP or AWS) and familiarity with observability (metrics, logging, alerting, SLOs/SLIs).
Experience:DevSecOpsSaaS
Skills:Growth mindsetStrong written communicationAbility to operate independentlyIncident response troubleshootingAutomation mindset
Tech Stack:KubernetesTerraformGoRubyCI/CDGitOpsAWSGCPMetricsLogsAlertingSLOsSLIs

Company Brief

GitLab
Provides a single application for the complete DevSecOps lifecycle, offering source code management, CI/CD, security, and collaboration tools to help teams deliver software faster and more securely.
Industry: Developer Tools
Company Size: Enterprise (1,001+ employees)
Revenue: USD 250M to 500M
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: San Francisco, United States
Founded: 2011
WebsiteLinkedIn