Senior Site Reliability Engineer
NVIDIA
India, Bengaluru
Workplace: RemoteFull timeFunction: DevOps, Cloud & InfrastructureExperience: 5+ yearsEducation: bachelorsSkills: ["Troubleshooting","Incident management","Root cause analysis","Blameless postmortems","Operational excellence"]Join the cloud service team to support, triage, and build the GeForce NOW cloud gaming platform with SRE practices that improve reliability and product quality. You’ll monitor production service health using metrics, logs, traces, and dashboards; lead incident response and blameless postmortems; and drive observability, automation, and self-service tooling. The role also includes operating Kubernetes-based services and partnering with service owners to improve SLOs.

