Senior Site Reliability Engineer

Camunda
Anywhere
Workplace: RemoteFull timeUSD 149,800 - 241,500 annuallyFunction: DevOps, Cloud & InfrastructureSkills: ["Incident response","Root cause analysis","Communication","Automation","Mentoring"]

Design and maintain Camunda’s Kubernetes-based multi-cloud platform, ensuring high availability, scalability, and fault tolerance. Build observability with monitoring and alerting tools (e.g., Prometheus/Grafana) and create runbooks and automation using a “you build it, you run it” model with on-call rotations. Collaborate with product, engineering, and support teams to ship improvements, respond to incidents with clear communication, and mentor others.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Camunda
Camunda
1 day ago

Senior Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 12 hours agoStatus: Live

Job Summary

Design and maintain Camunda’s Kubernetes-based multi-cloud platform, ensuring high availability, scalability, and fault tolerance. Build observability with monitoring and alerting tools (e.g., Prometheus/Grafana) and create runbooks and automation using a “you build it, you run it” model with on-call rotations. Collaborate with product, engineering, and support teams to ship improvements, respond to incidents with clear communication, and mentor others.
Location: Anywhere
Workplace: Remote
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Design and maintain a Kubernetes-based, multi-cloud platform architecture, ensuring availability, scalability, and fault tolerance.
  • •Implement and improve monitoring and alerting/observability so SREs and developers can understand system health and performance.
  • •Own systems end-to-end with a “you build it, you run it” approach, including on-call rotations, runbooks, and automation.
  • •Collaborate with product, engineering, product management, and support to define and ship features and improvements.
  • •Identify repetitive work and automate it away; share learning and mentor others on complex infrastructure challenges.

Pay and Benefits

Salary: USD 149,800 - 241,500 annually
Perks:Health InsuranceDentalVision401kRemote WorkLearning Budget

Key Requirements

  • •Deep hands-on experience building and maintaining Kubernetes clusters in production, including workloads, networking, and storage.
  • •Infrastructure as code expertise using tools like Terraform, with versioning, testing, and safe deployments.
  • •Demonstrated monitoring and observability experience using tools like Prometheus and Grafana, including alerting on what matters.
  • •Strong incident response and 3rd-level support skills, including root cause analysis and clear stakeholder communication under pressure.
  • •A passion for automation and raising the quality bar by building reliable, maintainable systems.
Experience:KubernetesCloudMulti-cloud
Skills:Incident responseRoot cause analysisCommunicationAutomationMentoring
Tech Stack:KubernetesTerraformPrometheusGrafanaArgoCDGitOpsAWSEKSGoogle Cloud PlatformGKEPythonGoTerrraform

Company Brief

Camunda
Provides an open-source and commercial platform for process orchestration and workflow automation, enabling enterprises to design, execute, and monitor business processes across people, systems, and AI agents.
Industry: SaaS
Company Size: Large (251 to 1,000 employees)
Growth: Scaleup
Funding: Series B
Headquarters: Berlin, Germany
Founded: 2008
Glassdoor
Glassdoor: 3.4
WebsiteLinkedInGlassdoor