Major Incident Manager

Sonar Source
Geneva
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureSkills: ["Change management","Automation mindset","Incident response","Prioritization","Reliability focus"]

Own major incident readiness and response by monitoring production infrastructure, triaging high-severity alerts, and managing error budgets using SLOs/SLIs. Build and maintain infrastructure- and policy-as-code to automate secure deployments and reduce operational toil. Strengthen DevSecOps security tooling in CI/CD and ensure observability across logging/metrics/tracing for fast diagnosis, then drive post-mortem learnings into automated runbooks to improve MTTR.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Sonar Source
Sonar Source
1 month ago

Major Incident Manager

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 6 hours agoStatus: Live

Job Summary

Own major incident readiness and response by monitoring production infrastructure, triaging high-severity alerts, and managing error budgets using SLOs/SLIs. Build and maintain infrastructure- and policy-as-code to automate secure deployments and reduce operational toil. Strengthen DevSecOps security tooling in CI/CD and ensure observability across logging/metrics/tracing for fast diagnosis, then drive post-mortem learnings into automated runbooks to improve MTTR.
Location: Geneva
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure

Key Responsibilities

  • •Monitor production infrastructure health, triage high-severity alerts, and manage error budgets using SLOs for daily prioritization.
  • •Develop and test infrastructure as code and policy as code to automate deployment, configuration, and security hardening while preventing configuration drift.
  • •Identify operational toil and implement automation for security patching, compliance checks, certificate rotations, and infrastructure maintenance.
  • •Maintain DevSecOps security tools in CI/CD pipelines and ensure robust logging, metrics, and tracing for immediate incident diagnostics.
  • •Participate in on-call and translate post-mortems into code-based preventative measures and automated runbook actions to reduce MTTR.

Key Requirements

  • •Proven experience provisioning and managing infrastructure with Terraform or CloudFormation (AWS), and/or configuration management tools like Ansible or Puppet.
  • •Hands-on experience with a major cloud provider (AWS, GCP, Azure) or large-scale internal/private cloud infrastructure.
  • •Ability to define, measure, and report SLOs/SLIs for critical services, including error budget monitoring.
  • •Experience with observability stacks such as Prometheus/Grafana, ELK/EFK, or vendor tools like Datadog/Splunk.
  • •Strong networking knowledge (TCP/IP, DNS, load balancing, firewalls, proxies) to debug connectivity and latency issues.
Experience:DevSecOpsInfrastructure as CodeObservabilityCloudSecurity
Skills:Change managementAutomation mindsetIncident responsePrioritizationReliability focus
Tech Stack:PythonGoTerraformCloudFormationAWSGCPAzureAnsiblePuppetPrometheusGrafanaELKEFKDatadogSplunkCI/CDDevSecOpsSLOSLIHashiCorp Vault

Company Brief

Sonar Source
Builds static code analysis and continuous inspection tools (SonarQube, SonarCloud, SonarLint) that identify bugs, vulnerabilities, and code smells across multiple languages to help teams improve code quality and maintainability.
Industry: Developer Tools
Company Size: Large (251 to 1,000 employees)
Growth: Established Company
Headquarters: Geneva, Switzerland
Founded: 2008
WebsiteLinkedIn