Site Reliability Engineer (Onsite Hybrid)
NTT
Plano
Workplace: HybridFull timeFunction: DevOps, Cloud & InfrastructureExperience: 5+ yearsSkills: ["Incident response","Root-cause analysis","Troubleshooting","Alerting strategy"]Own end-to-end reliability for production systems by driving observability with New Relic (APM, dashboards, alerting), defining SLIs/SLOs, and leading incident response with root-cause analysis and post-mortems. Administer GitHub Enterprise and build CI/CD reliability for Java/.NET applications using GitHub Actions. Troubleshoot across application, infrastructure, and pipeline layers, continuously improving performance and leveraging AI/automation in SRE workflows.

