Site Reliability Engineer
NTT
Guadalajara
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureExperience: 6-8 yearsSkills: ["Communication","Incident response","Troubleshooting","Automation","Production support"]Own and improve the reliability of microservice-based cloud services, driving incident response, production support, and blameless postmortems. You’ll manage Kubernetes workloads and distributed systems, implement Infrastructure as Code with Terraform, and enhance GitOps and CI/CD workflows using ArgoCD and related tools. Work across AWS and Azure environments, improve observability and automation, and ensure scalability, resilience, and disaster recovery readiness.

