Senior Site Reliability Engineer

GovCIO
United States
Workplace: HybridFull timeUSD 210,000 - 230,000 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 12+ yearsEducation: bachelorsSkills: ["Problem-solving","Analytical ability","Communication","Ability to work independently","Customer-focused mindset"]

Design, implement, and maintain highly available, scalable, and resilient infrastructure across multi-cloud environments. Own infrastructure automation using IaC (Terraform) and configuration management (Ansible), build CI/CD and GitOps workflows, and implement monitoring, SLO/SLI practices, and incident/on-call response. Lead disaster recovery and business continuity planning, drive chaos engineering experiments, and collaborate with development teams to continuously improve reliability and performance.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
GovCIO
GovCIO
1 day ago

Senior Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 4 hours agoStatus: Live

Job Summary

Design, implement, and maintain highly available, scalable, and resilient infrastructure across multi-cloud environments. Own infrastructure automation using IaC (Terraform) and configuration management (Ansible), build CI/CD and GitOps workflows, and implement monitoring, SLO/SLI practices, and incident/on-call response. Lead disaster recovery and business continuity planning, drive chaos engineering experiments, and collaborate with development teams to continuously improve reliability and performance.
Location: United States
Workplace: Hybrid
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Design, deploy, and manage cloud infrastructure using Infrastructure as Code principles.
  • •Develop Terraform modules, Ansible playbooks, and self-service tooling for development teams.
  • •Implement monitoring, logging, and alerting; define and monitor SLOs/SLIs; conduct capacity planning and performance tuning.
  • •Lead disaster recovery and business continuity strategies, including backup/restore, failover automation, and RTO/RPO documentation.
  • •Participate in on-call rotation and incident response, collaborate on architecture decisions, and mentor teammates in SRE practices.

Pay and Benefits

Salary: USD 210,000 - 230,000 annually

Key Requirements

  • •Bachelor’s degree with 12+ years of experience.
  • •Active Secret clearance with the ability to obtain and hold DEA suitability.
  • •3+ years hands-on experience with AWS and Azure.
  • •Expert proficiency with Terraform and strong experience with Ansible.
  • •Proficiency with scripting (Python, Bash, or PowerShell) and containerization (Docker, Kubernetes), plus strong Git/GitHub workflow knowledge.
Experience:12+ yearsMulti-cloudDevOpsPlatform engineeringInfrastructure automationSRE
Education:Bachelor's
Skills:Problem-solvingAnalytical abilityCommunicationAbility to work independentlyCustomer-focused mindset
Certifications:AWS Certified Solutions ArchitectAWS Certified SysOps AdministratorAzure AdministratorAzure Solutions Architect certificationCertified Kubernetes Administrator (CKA)HashiCorp Certified: Terraform AssociateGitHub CertifiedGitHub Enterprise
Languages:En-us
Tech Stack:AWSAzureTerraformAnsiblePythonBashPowerShellDockerKubernetesGitGitHubGitOpsPrometheusGrafanaELK StackGitHub ActionsJenkinsGitLab CIAzure DevOpsEC2

Eligibility

Security Clearance:SecretDEA suitability

Company Brief

GovCIO
Provides IT modernization, cloud, cybersecurity, data analytics, and managed services to U.S. federal civilian and defense agencies, delivering technology solutions and mission support to improve government operations and security.
Industry: Professional Services
Growth: Established Company
Headquarters: Herndon, United States
WebsiteLinkedIn