Site Reliability Engineer

Delinea
Manila
Full timeFunction: DevOps, Cloud & InfrastructureExperience: 5+ yearsSkills: ["Ownership","Customer-first mindset","Incident communication","Problem-solving","Collaboration"]

Own the availability, performance, and reliability of production SaaS applications running on Azure across multiple regions. Lead troubleshooting for AKS, deployments, networking, and autoscaling issues, and drive incident response during an on-call rotation. Improve disaster recovery, failover, and incident management processes for multi-region deployments by building automation and monitoring. Write RCAs and partner cross-functionally to strengthen observability and performance best practices.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Delinea
Delinea
2 days ago

Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 2 hours agoStatus: Live

Job Summary

Own the availability, performance, and reliability of production SaaS applications running on Azure across multiple regions. Lead troubleshooting for AKS, deployments, networking, and autoscaling issues, and drive incident response during an on-call rotation. Improve disaster recovery, failover, and incident management processes for multi-region deployments by building automation and monitoring. Write RCAs and partner cross-functionally to strengthen observability and performance best practices.
Location: Manila
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Own availability and performance for Azure-hosted production SaaS applications across multiple geographic regions.
  • •Troubleshoot and resolve cloud infrastructure and application issues including AKS pod/node failures, deployment rollbacks, ingress/networking problems, and resource/autoscaling issues.
  • •Participate in on-call rotation and lead incident response from detection through resolution, minimizing customer impact.
  • •Improve disaster recovery, failover, and incident management processes across multi-region deployments.
  • •Build and maintain automation and monitoring tools, and write RCAs to identify root causes and drive preventive actions to closure.

Pay and Benefits

Perks:Health InsurancePensionLife InsuranceEmployee AssistancePaid Holidays

Key Requirements

  • •5+ years of experience in Site Reliability Engineering, DevOps, or Cloud Administration with ownership of production systems.
  • •Hands-on Azure administration including AKS (Kubernetes), core Azure services, cloud networking, and cloud security fundamentals.
  • •Experience with monitoring, logging, and alerting (e.g., Datadog, Azure Monitor, ELK) and troubleshooting using logs and traces with Datadog APM.
  • •Familiarity with networking fundamentals including firewalls, load balancers, VPNs, DNS, and routing.
  • •Experience with automation and scripting using PowerShell, Python, or similar, plus backup and disaster recovery for geo-redundant/multi-region deployments.
Experience:5+ yearsSaaSDevOpsMulti-cloudMulti-region
Skills:OwnershipCustomer-first mindsetIncident communicationProblem-solvingCollaboration
Tech Stack:AzureAKSApp ServiceRedisSQLService BusKubernetesDatadogAzure MonitorELK stackDatadog APMPowerShellPythonAWSAzure DevOpsTerraformARM templatesCI/CDLoad balancersVPNs

Company Brief

Delinea
Provides privileged access management and identity security solutions that centralize authorization for human and machine identities, securing hybrid-cloud infrastructure, applications, and data for enterprises worldwide.
Industry: Cybersecurity
Company Size: Enterprise (1,001+ employees)
Revenue: USD 250M to 500M
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Private Equity Backed
Headquarters: San Francisco, United States
Founded: 2004
Glassdoor
Glassdoor: 3.6
WebsiteLinkedInGlassdoor