Senior Site Reliability Engineer, AI Platform

Algolia
Paris
Workplace: OnsiteFull timeEUR 69,768 - 96,900 annuallyFunction: DevOps, Cloud & InfrastructureSkills: ["Ownership","Problem-solving","Communication","Automation mindset","Mentoring"]

Own and evolve production infrastructure for AI-related workloads, driving reliable delivery at scale. Design and operate highly available Kubernetes-based platforms, improving reliability through SLOs, observability, capacity planning, and production guardrails. Lead complex incident investigations and turn insights into durable architectural improvements across networking, databases, and service communication. Build CI/CD and progressive delivery automation, drive cloud efficiency and FinOps, and mentor engineers.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Algolia
Algolia
22 hours ago

Senior Site Reliability Engineer, AI Platform

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 13 minutes agoStatus: Live

Job Summary

Own and evolve production infrastructure for AI-related workloads, driving reliable delivery at scale. Design and operate highly available Kubernetes-based platforms, improving reliability through SLOs, observability, capacity planning, and production guardrails. Lead complex incident investigations and turn insights into durable architectural improvements across networking, databases, and service communication. Build CI/CD and progressive delivery automation, drive cloud efficiency and FinOps, and mentor engineers.
Location: Paris
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Own and evolve production infrastructure supporting AI-related workloads and services at scale.
  • •Design and operate highly available Kubernetes-based platforms.
  • •Drive reliability through SLOs, observability, capacity planning, and production guardrails.
  • •Lead complex production investigations and translate findings into durable architectural improvements.
  • •Improve shared infrastructure across networking, databases, service communication, and compute, while building better CI/CD and automation.

Pay and Benefits

Salary: EUR 69,768 - 96,900 annually

Key Requirements

  • •Strong hands-on production experience with at least one major cloud provider: GCP, AWS, or Azure.
  • •Strong experience designing and operating Kubernetes and cloud-native production systems at scale.
  • •Strong understanding of distributed systems, networking, and reliability engineering.
  • •Experience operating business-critical systems with strong availability, scalability, and operational requirements.
  • •Ability to independently own ambiguous, cross-team technical problems and drive measurable outcomes.
Experience:AI/MLCloud-native
Skills:OwnershipProblem-solvingCommunicationAutomation mindsetMentoring
Languages:English
Tech Stack:KubernetesGCPAWSAzureCI/CDCI CDProgressive deliveryNetworkingDatabasesObservabilityFinOpsCloud infrastructureOn-callIncident responseGoPythonGPUs

Company Brief

Algolia
Provides a hosted search and discovery API that enables developers and product teams to build fast, relevant search and discovery experiences across websites and applications with features like instant search, personalization, and analytics.
Industry: API Platforms
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series D
Headquarters: San Francisco, United States
Founded: 2012
WebsiteLinkedIn