Director, Engineering Operations and Site Reliability Engineering - Datacenter Server Systems
NVIDIA
Santa Clara
Workplace: OnsiteFull timeUSD 292,000 - 442,750 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 12+ yearsEducation: bachelorsSkills: ["Leadership","Communication","Mentoring","Strategic thinking","Team building"]Lead and scale NVIDIA’s engineering operations and site reliability for datacenter server systems. Drive fleet operations, incident response, and reliability metrics; build automation, telemetry, and dashboards; partner across hardware, software, networking, validation, and infrastructure teams to ensure high availability and fast development velocity.

