Infrastructure engineer (UK)

Writer
London
Workplace: HybridFull timeFunction: DevOps, Cloud & InfrastructureSkills: ["Communication","Collaboration","Problem-solving","Ownership","Decision-making"]

Own the reliability, performance, and efficiency of core services powering AI-powered enterprise workflows. Build and automate scalable, fault-tolerant infrastructure across AWS (preferred), GCP, and Azure using Kubernetes, Helm, and Terraform, while leading incident response, post-mortems, and root-cause analysis. Use AI-assisted tooling in your daily workflow to investigate incidents, draft changes, write runbooks, and review PRs, collaborating closely with product and security teams.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Writer
Writer
2 days ago

Infrastructure engineer (UK)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 17 minutes agoStatus: Live
Reposted: similar role first listed 3 months ago

Job Summary

Own the reliability, performance, and efficiency of core services powering AI-powered enterprise workflows. Build and automate scalable, fault-tolerant infrastructure across AWS (preferred), GCP, and Azure using Kubernetes, Helm, and Terraform, while leading incident response, post-mortems, and root-cause analysis. Use AI-assisted tooling in your daily workflow to investigate incidents, draft changes, write runbooks, and review PRs, collaborating closely with product and security teams.
Location: London
Workplace: Hybrid
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Own end-to-end reliability, performance, and efficiency of core services by defining/upholding SLOs and error budgets and standing behind outcome metrics.
  • •Build scalable, fault-tolerant infrastructure across AWS (preferred), GCP, and Azure using Kubernetes, Helm, and Terraform (or Pulumi).
  • •Lead incident response, post-mortems, and root-cause analyses, applying learnings to prevent repeat failures.
  • •Automate operational tasks and infrastructure management, reducing toil and improving on-call and release processes.
  • •Use AI-assisted workflows (agents) to investigate incidents, draft Terraform/Helm changes, write runbooks, scaffold tooling, and review PRs.

Pay and Benefits

Perks:Paid LeaveHealth InsuranceDentalPensionEquityLearning BudgetWellness Stipend

Key Requirements

  • •5+ years of experience in infrastructure engineering or DevOps building and operating large-scale, high-availability production systems.
  • •Production experience with containerisation and at least Helm and Terraform (or Pulumi) on a major cloud (AWS preferred).
  • •Strong automation skills in Python or Go for operational tasks and infrastructure management.
  • •Experience using AI/agentic tooling in your daily workflow (e.g., Claude Code, Droid, Codex) to investigate incidents and drive changes.
  • •Demonstrated ability to reason from failure modes, challenge best practices, and propose tradeoff-aware reliability improvements.
Experience:High-availability production systemsContainerisationEnterprise softwareCloud infrastructureAI in operations
Skills:CommunicationCollaborationProblem-solvingOwnershipDecision-making
Tech Stack:AWSGCPAzureKubernetesHelmTerraformPulumiPythonGoPrometheusGrafanaELKClaude CodeDroidCodex

Company Brief

Writer
Provides an AI writing platform for enterprises that helps teams create on-brand, high-quality content at scale using customizable style guides, real-time suggestions, and governance controls.
Industry: AI & Machine Learning
Company Size: Medium (51 to 250 employees)
Growth: Growth Stage Startup
Funding: Series B
Headquarters: New York, United States
Founded: 2019
WebsiteLinkedIn