Site Reliability Engineer
Austin
Workplace: HybridFull timeFunction: DevOps, Cloud & InfrastructureExperience: 5+ yearsSkills: ["Incident management","Observability","Automation","Reliability engineering","Continuous improvement"]Build and operate a product-specific SRE team to deliver high availability and strong performance for a public-cloud telecommunications solution. Own service reliability by defining SLOs/SLIs and error budgets, running 24/7 on-call and deep-dive incident troubleshooting, and driving CI/CD and infrastructure automation with Terraform, Ansible, Kubernetes, and GitLab. Implement observability with Datadog, lead postmortems, and collaborate with cloud security to improve access controls and respond to vulnerabilities.
Loading
Loading job details...
Preparing the role view and application actions.

