Site Reliability Engineer, AI Platform

Algolia
Paris
Workplace: OnsiteFull timeEUR 69,768 - 96,900 annuallyFunction: DevOps, Cloud & InfrastructureSkills: ["Automation mindset","Problem-solving","Ownership","Troubleshooting","Collaboration","Attention to reliability"]

Build and operate production infrastructure for AI workloads, running highly available Kubernetes-based platforms. Improve reliability using SLOs, observability, alerting, and capacity management, and turn incident learnings into durable fixes. Collaborate across networking, databases, compute, and service infrastructure while enhancing CI/CD pipelines, deployment automation, and developer experience through Infrastructure as Code.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Algolia
Algolia
22 hours ago

Site Reliability Engineer, AI Platform

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 16 minutes agoStatus: Live

Job Summary

Build and operate production infrastructure for AI workloads, running highly available Kubernetes-based platforms. Improve reliability using SLOs, observability, alerting, and capacity management, and turn incident learnings into durable fixes. Collaborate across networking, databases, compute, and service infrastructure while enhancing CI/CD pipelines, deployment automation, and developer experience through Infrastructure as Code.
Location: Paris
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure

Key Responsibilities

  • •Build and operate production infrastructure supporting AI-related workloads and services.
  • •Operate and improve highly available Kubernetes-based platforms.
  • •Improve reliability through SLOs, observability, alerting, and capacity management.
  • •Investigate production issues and turn findings into durable fixes and improvements.
  • •Improve CI/CD pipelines and deployment automation, using Infrastructure as Code; participate in on-call and incident response.

Pay and Benefits

Salary: EUR 69,768 - 96,900 annually

Key Requirements

  • •Strong hands-on Kubernetes knowledge, including workloads, resource management, and production operations.
  • •Experience with Infrastructure as Code and the lifecycle of cloud infrastructure.
  • •Experience building and operating CI/CD pipelines and automated deployment workflows.
  • •Hands-on experience with at least one major cloud provider (GCP, AWS, or Azure).
  • •Good understanding of networking, distributed systems, reliability engineering, and monitoring/observability for troubleshooting.
Skills:Automation mindsetProblem-solvingOwnershipTroubleshootingCollaborationAttention to reliability
Languages:English
Tech Stack:KubernetesCI/CDInfrastructure as CodeGCPAWSAzureSLOsObservabilityAlertingFinOpsNetworkingDatabasesCloud infrastructureDeployment automation

Company Brief

Algolia
Provides a hosted search and discovery API that enables developers and product teams to build fast, relevant search and discovery experiences across websites and applications with features like instant search, personalization, and analytics.
Industry: API Platforms
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series D
Headquarters: San Francisco, United States
Founded: 2012
WebsiteLinkedIn