DevOps Engineer

Encord
London
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureExperience: 4+ yearsSkills: ["Automation","Collaboration","Reliability mindset"]

Embedded in platform engineering, you’ll own and improve CI/CD and deployment pipelines while designing, deploying, and operating cloud infrastructure on GCP and AWS. You’ll drive automation to reduce toil, optimize performance and capacity for large-scale data workloads, and ensure reliability through SLIs/SLOs, alerting, runbooks, and blameless postmortems. You’ll also raise observability standards using tracing, logging, and metrics tools so services are production-ready.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Encord
Encord
1 day ago

DevOps Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 6 hours agoStatus: Live

Job Summary

Embedded in platform engineering, you’ll own and improve CI/CD and deployment pipelines while designing, deploying, and operating cloud infrastructure on GCP and AWS. You’ll drive automation to reduce toil, optimize performance and capacity for large-scale data workloads, and ensure reliability through SLIs/SLOs, alerting, runbooks, and blameless postmortems. You’ll also raise observability standards using tracing, logging, and metrics tools so services are production-ready.
Location: London
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Own and continuously improve deployment pipelines, partnering with developers to review infrastructure changes and streamline releases.
  • •Design, deploy, and maintain cloud infrastructure on GCP and AWS, including Kubernetes clusters, networking, and storage using infrastructure-as-code.
  • •Build and guide automation and internal tooling to eliminate manual toil and improve developer productivity across squads.
  • •Optimize services for performance and capacity, including capacity planning and establishing performance benchmarks and expectations.
  • •Define SLIs/SLOs/SLAs, build alerting and runbooks, lead incident response and blameless postmortems, and ensure production observability using traces, logs, and metrics.
Travel: Medium travel

Pay and Benefits

Perks:EquityPaid LeaveLearning Budget

Key Requirements

  • •4+ years of hands-on DevOps, platform engineering, or SRE experience in a production environment.
  • •Strong experience building and maintaining CI/CD pipelines and deployment automation at scale.
  • •Proven experience with infrastructure-as-code tools such as Terraform and Pulumi, plus configuration management.
  • •Hands-on experience with Kubernetes and containerised workloads in cloud environments (GCP and/or AWS).
  • •Experience defining reliability and observability practices, including metrics, logs, traces, and alerting.
Experience:4+ years
Skills:AutomationCollaborationReliability mindset
Tech Stack:PythonTypeScriptReactKubernetesGCPAWSTerraformPulumiPrometheusGrafanaOpenTelemetryGCP DashboardsPyTorchCUDARay

Company Brief

Encord
Provides a computer vision data platform for labeling, dataset management, and model evaluation, enabling teams to build, validate, and monitor ML models with collaboration and tooling for high-quality annotated data.
Industry: Data Infrastructure
Company Size: Medium (51 to 250 employees)
Growth: Growth Stage Startup
Funding: Series A
Headquarters: London, United Kingdom
Founded: 2019
WebsiteLinkedIn