Site Reliability Engineer, IaaS

Algolia
Paris
Full timeFunction: DevOps, Cloud & InfrastructureSkills: ["Problem-solving","Communication","Teamwork","Reliability focus","Good judgement"]

Build and evolve Algolia’s production infrastructure for a unified cloud and Kubernetes platform. As an IaaS Site Reliability Engineer (P3), you’ll create cloud baseline capabilities (identity, networking, security, inventory), implement infrastructure-as-code and automation, and deliver reliable cluster lifecycle operations. You’ll improve observability and alerting, reduce manual work and drift with GitOps/testing standards, and help investigate incidents through on-call while partnering with Infrastructure, Security, FinOps, and engineering teams.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Algolia
Algolia
4 days ago

Site Reliability Engineer, IaaS

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 12 hours agoStatus: Live

Job Summary

Build and evolve Algolia’s production infrastructure for a unified cloud and Kubernetes platform. As an IaaS Site Reliability Engineer (P3), you’ll create cloud baseline capabilities (identity, networking, security, inventory), implement infrastructure-as-code and automation, and deliver reliable cluster lifecycle operations. You’ll improve observability and alerting, reduce manual work and drift with GitOps/testing standards, and help investigate incidents through on-call while partnering with Infrastructure, Security, FinOps, and engineering teams.
Location: Paris
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Build and improve Cloud Baseline capabilities including identity and access, networking, security, resource inventory, tagging, and auditability.
  • •Develop and maintain infrastructure as code and automation for cloud and Kubernetes.
  • •Deliver reliable, repeatable cloud and cluster lifecycle operations.
  • •Reduce manual work and configuration drift using automation, testing, GitOps practices, and standardization.
  • •Improve observability, monitoring, alerting, capacity management, and operational documentation; investigate production issues and participate in on-call.

Key Requirements

  • •Hands-on production knowledge of AWS or GCP.
  • •Practical Kubernetes experience and interest in operating it in production.
  • •Familiarity with infrastructure as code, ideally Terraform.
  • •Programming or scripting skills in Python, Go, or an equivalent language.
  • •Strong Linux and networking fundamentals with a reliability and automation mindset.
Experience:Production operationsCloud migrationPlatform engineering
Skills:Problem-solvingCommunicationTeamworkReliability focusGood judgement
Languages:English
Tech Stack:AWSGCPKubernetesTerraformPythonGoLinuxNetworkingGitOpsArgo CDHelmOPAKyverno

Company Brief

Algolia
Provides a hosted search and discovery API that enables developers and product teams to build fast, relevant search and discovery experiences across websites and applications with features like instant search, personalization, and analytics.
Industry: API Platforms
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series D
Headquarters: San Francisco, United States
Founded: 2012
WebsiteLinkedIn