Site Reliability Engineer

Kong
Milan
Workplace: HybridFull timeFunction: DevOps, Cloud & InfrastructureSkills: ["Incident resolution","Automation","Collaboration","Reliability focus","Observability mindset"]

Build and maintain Kong’s core infrastructure as code to power high-reliability cloud services. Implement monitoring, logging, and alerting to exceed 99.99% uptime, and lead incident resolution with blameless post-mortems. Automate operational tasks, collaborate with developers to bake in reliability and scalability, and support capacity planning, disaster recovery drills, and security hardening while participating in an on-call rotation.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Kong
Kong
1 month ago

Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 10 hours agoStatus: Live

Job Summary

Build and maintain Kong’s core infrastructure as code to power high-reliability cloud services. Implement monitoring, logging, and alerting to exceed 99.99% uptime, and lead incident resolution with blameless post-mortems. Automate operational tasks, collaborate with developers to bake in reliability and scalability, and support capacity planning, disaster recovery drills, and security hardening while participating in an on-call rotation.
Location: Milan
Workplace: Hybrid
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure

Key Responsibilities

  • •Build and maintain core infrastructure as code using tools like Terraform and Ansible.
  • •Implement monitoring, logging, and alerting to ensure services meet and exceed 99.99% uptime.
  • •Resolve production incidents through systematic debugging and run blameless post-mortems to prevent recurrence.
  • •Write automation to reduce operational toil and improve system efficiency for engineering self-service.
  • •Contribute to capacity planning, disaster recovery drills, and security hardening, while participating in on-call rotation.

Key Requirements

  • •Experience operating production workloads on a major cloud provider (AWS, GCP, Azure).
  • •Proficiency in at least one programming or scripting language such as Golang, Python, or Bash.
  • •Hands-on experience with containerization and orchestration technologies like Docker and Kubernetes.
  • •Knowledge of Infrastructure as Code principles and tools (Terraform is a plus).
  • •Familiarity with CI/CD concepts and pipeline tools such as GitLab CI or Jenkins, plus observability stacks like Prometheus, Grafana, and ELK.
Skills:Incident resolutionAutomationCollaborationReliability focusObservability mindset
Tech Stack:TerraformAnsibleAWSGCPAzureGolangPythonBashDockerKubernetesCI/CDGitLab CIJenkinsPrometheusGrafanaELK

Company Brief

Kong
Provides a cloud-native API and service connectivity platform (Kong Gateway, Konnect, Kuma, Insomnia) that secures, manages, and observes APIs and microservices for enterprises across cloud and hybrid environments.
Industry: API Platforms
Company Size: Large (251 to 1,000 employees)
Revenue: USD 100M to 250M
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2017
Glassdoor
Glassdoor: 3.9
WebsiteLinkedInGlassdoor