Site Reliability Engineer
NTT
Guadalajara
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureSkills: ["Communication","Incident response","Automation mindset","Production support","Blameless postmortems"]Own and improve reliability for cloud-based, microservice platforms, partnering with engineering teams to enhance incident response, uptime, and performance. Participate in on-call, drive blameless postmortems, and track corrective actions. Build and maintain infrastructure as code with Terraform and GitOps/deployment workflows using Atlantis, ArgoCD, and CI/CD pipelines across AWS and Azure. Manage Kubernetes and containers, improve observability, and support recovery readiness.

