Senior Staff Site Reliability Operations Technical Lead

NVIDIA
Durham
Workplace: OnsiteFull timeUSD 184,000 - 264,500 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 12+ yearsEducation: bachelorsSkills: ["Technical leadership","Executive-level communication","Root cause analysis","Mentorship","Diagnostic rigor"]

Serve as a senior technical individual contributor for Site Reliability Operations at NVIDIA’s Durham site, owning day-to-day service delivery, incident response, queue health, SLAs, and service quality. Act as Tier 3 escalation for identity, messaging, endpoints, compute, and datacenter/lab hardware; drive root-cause and permanent fixes. Lead site SRO engineers on standards and runbooks, build automation for diagnostics and remediation, and represent local needs in regional and global IT initiatives.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
1 day ago

Senior Staff Site Reliability Operations Technical Lead

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 1 hour agoStatus: Live

Job Summary

Serve as a senior technical individual contributor for Site Reliability Operations at NVIDIA’s Durham site, owning day-to-day service delivery, incident response, queue health, SLAs, and service quality. Act as Tier 3 escalation for identity, messaging, endpoints, compute, and datacenter/lab hardware; drive root-cause and permanent fixes. Lead site SRO engineers on standards and runbooks, build automation for diagnostics and remediation, and represent local needs in regional and global IT initiatives.
Location: Durham
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Own day-to-day site operations including incidents, requests, critical issues, and support coverage with accountability for queue health, SLA attainment, backlog, and service quality.
  • •Serve as Tier 3 escalation for identity, messaging, compute, and endpoints, driving root cause and permanent fixes through diagnostic evidence.
  • •Own endpoint compliance, vulnerability remediation, patch management, and hardening, and partner with InfoSec on incident response and privileged access.
  • •Drive critical issues into global platform teams and vendors with reproduction cases and diagnostic evidence to reach committed fixes.
  • •Provide technical leadership to site SRO engineers by setting standards, reviewing work, mentoring diagnostics, and maintaining the site knowledge base and runbook library.

Pay and Benefits

Salary: USD 184,000 - 264,500 annually
Equity and Bonus:Equity

Key Requirements

  • •12+ years in enterprise support engineering, infrastructure, or end user services, including 5+ years in a senior/lead/escalation-tier role in a multi-site environment.
  • •Deep hands-on expertise across Active Directory and hybrid Entra ID, Exchange hybrid, Windows and Linux server, virtualization, enterprise storage, and datacenter hardware.
  • •Strong endpoint management experience (Intune, Autopilot, MECM/SCCM, Jamf or equivalent), Windows 11, Microsoft 365, and endpoint security/vulnerability remediation.
  • •Database operations support and networking fundamentals including DNS, DHCP, VLAN, wireless, firewall policy, and switch-level troubleshooting.
  • •Scripting/automation in Python, PowerShell, or Bash for real support problems and experience with ServiceNow or similar ITSM; bachelor’s degree in CS/IS or equivalent experience.
Experience:12+ yearsEnterprise supportInfrastructureEnd user servicesMulti-site environmentsDatacenterEndpoint management
Education:Bachelor's in Computer Science, Information Systems, or related field
Skills:Technical leadershipExecutive-level communicationRoot cause analysisMentorshipDiagnostic rigor
Tech Stack:PowerShellPythonBashActive DirectoryEntra IDGPOKerberosLDAPSSOMFAExchange hybridSMTP relayWindowsLinuxMacOSVirtualizationStorageM365TeamsIntune

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor