Senior Site Reliability Engineer
Austin
Workplace: HybridFull timeFunction: DevOps, Cloud & InfrastructureExperience: 5+ yearsSkills: ["Hands-on leadership","Incident response","Reliability engineering","Cross-team collaboration","Systems thinking"]Own reliability for all 2K player-facing infrastructure, from multi-cloud and hybrid platforms to Kubernetes and progressive delivery. Build and run Terraform/Pulumi + GitOps (ArgoCD/Flux) infrastructure, manage EKS/GKE clusters with networking and autoscaling, and deliver SLI/SLO-driven observability using Prometheus, Grafana, Datadog, and OpenTelemetry. Lead incident response, post-mortems, and chaos engineering while hardening CI/CD and embedding security at the platform layer.

