Site Reliability Engineer (Europe)

ArangoDB
India
Workplace: RemoteFull timeFunction: DevOps, Cloud & InfrastructureSkills: ["Self-organization","Communication","Troubleshooting","Collaboration","Automation"]

Design, implement, and maintain cloud infrastructure for distributed database systems running on Kubernetes across AWS and Google Cloud. Improve scalability, reliability, and performance through automation, CI/CD pipeline optimization, and expanded observability with monitoring, logging, and alerting. Troubleshoot complex issues across network, OS, and cloud layers, support disaster recovery and high availability, and write production-grade automation code in Golang while collaborating with developers and cross-functional teams.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
ArangoDB
ArangoDB
2 days ago

Site Reliability Engineer (Europe)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 17 hours agoStatus: Live

Job Summary

Design, implement, and maintain cloud infrastructure for distributed database systems running on Kubernetes across AWS and Google Cloud. Improve scalability, reliability, and performance through automation, CI/CD pipeline optimization, and expanded observability with monitoring, logging, and alerting. Troubleshoot complex issues across network, OS, and cloud layers, support disaster recovery and high availability, and write production-grade automation code in Golang while collaborating with developers and cross-functional teams.
Location: India
Workplace: Remote
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure

Key Responsibilities

  • •Design, implement, and maintain cloud infrastructure on AWS and Google Cloud.
  • •Ensure scalability, performance, and reliability of Kubernetes-based distributed database systems.
  • •Collaborate with developers to write production-grade Golang code to automate infrastructure management and operations.
  • •Optimize and automate CI/CD pipelines, deployments, and monitoring systems for production.
  • •Implement disaster recovery, high availability, and fault tolerance, and participate in on-call rotations to respond to incidents.

Key Requirements

  • •SRE or DevOps engineering experience in cloud-native environments, with strong communication and self-organized remote work style.
  • •Proficiency with AWS and Google Cloud, Linux internals, and containerization/orchestration (Docker, Kubernetes at scale).
  • •Experience improving CI/CD pipelines and observability using tools such as Jenkins, CircleCI, Prometheus, Grafana, and ELK, plus Git version control.
  • •Strong understanding of networking and security best practices with the ability to troubleshoot complex infrastructure issues systematically.
  • •Programming proficiency in Golang or Python, with willingness to learn Golang if needed.
Experience:Cloud-nativeKubernetesDistributed systems
Skills:Self-organizationCommunicationTroubleshootingCollaborationAutomation
Tech Stack:AWSGoogle CloudGolangPythonLinuxDockerKubernetesJenkinsCircleCIPrometheusGrafanaELK stackGitTerraformGitOpsBashCI/CDInfrastructure-as-Code (IaC)

Company Brief

ArangoDB
ArangoDB builds a multi-model database combining graph, document, and key-value models with a unified query language and managed cloud (Oasis), enabling scalable graph analytics, search, and ML use cases for enterprises.
Industry: Data Infrastructure
Company Size: Medium (51 to 250 employees)
Growth: Scaleup
Funding: Private Equity Backed
Headquarters: San Francisco, United States
Founded: 2015
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor