Cloud Site Reliability Senior Engineer

Barracuda Networks
Bengaluru
Full timeFunction: DevOps, Cloud & InfrastructureExperience: 5+ yearsEducation: bachelorsSkills: ["Communication","Collaboration","Troubleshooting","Problem-solving","Independent work"]

Lead reliability improvements for production cloud services within a centralized Cloud Operations team. Drive incident response, RCA, and preventive actions to improve availability and operational excellence, partnering with Engineering, Cloud Operations, and SRE teams. Strengthen observability, alerting, and operational insights, and support secure, compliant, cost-conscious cloud operations across Azure, AWS, Terraform-managed infrastructure, and containerized platforms.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Barracuda Networks
Barracuda Networks
5 days ago

Cloud Site Reliability Senior Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 3 hours agoStatus: Live

Job Summary

Lead reliability improvements for production cloud services within a centralized Cloud Operations team. Drive incident response, RCA, and preventive actions to improve availability and operational excellence, partnering with Engineering, Cloud Operations, and SRE teams. Strengthen observability, alerting, and operational insights, and support secure, compliant, cost-conscious cloud operations across Azure, AWS, Terraform-managed infrastructure, and containerized platforms.
Location: Bengaluru
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Improve reliability, availability, and operational excellence of production services through shared team ownership.
  • •Lead troubleshooting, incident response, RCA, and preventive improvements across cloud services and supporting platforms.
  • •Partner with Engineering, Cloud Operations, and SRE teams to improve release readiness, deployment practices, and operational standards.
  • •Support cloud platform improvements, automation, and resilience initiatives to reduce toil and improve service health.
  • •Strengthen observability and operational insights (alerting, monitoring) to help teams respond faster and improve service reliability.

Key Requirements

  • •Bachelor’s degree in computer science engineering, Information Technology, or equivalent.
  • •5+ years of progressive experience in Site Reliability Engineering, Cloud Operations, DevOps, or Platform Engineering.
  • •Strong Linux/Unix command-line administration, troubleshooting, package management, and systems operations skills.
  • •Experience managing production cloud infrastructure across Azure and/or AWS environments.
  • •Hands-on Infrastructure as Code experience, preferably with Terraform.
Experience:5+ yearsSite reliability engineeringCloud operationsDevopsPlatform engineering
Education:Bachelor's
Skills:CommunicationCollaborationTroubleshootingProblem-solvingIndependent work
Tech Stack:AzureAWSTerraformAKSInfrastructure as CodePythonBashGoYAMLGenerative AIAzure DevOpsArgoCDJenkinsGitHub ActionsDockerAzure Container RegistryAnsiblePuppetChefPagerDuty

Company Brief

Barracuda Networks
Provides cloud-enabled security and data protection solutions including email protection, network and application security, and backup and recovery services for businesses, service providers, and government organizations worldwide.
Industry: Cybersecurity
Company Size: Enterprise (1,001+ employees)
Revenue: USD 250M to 500M
Growth: Established Company
Funding: Private Equity Backed
Headquarters: Campbell, United States
Founded: 2003
WebsiteLinkedIn