Senior Staff Site Reliability Operations

NVIDIA
Seattle
Workplace: OnsiteFull timeUSD 184,000 - 264,500 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 12+ yearsEducation: bachelorsSkills: ["Technical leadership","Executive communication","Root cause focus","Mentorship","Diagnostic rigor"]

Own day-to-day site reliability operations for NVIDIA’s Seattle, WA location, leading incidents, escalations, and service delivery quality. Serve as Tier 3 escalation across identity, messaging, endpoints, compute, and datacenter/lab hardware, driving root-cause analysis and permanent fixes. Lead standards and mentorship for site SRO engineers, manage compliance and patching, build automation with PowerShell/Python/Bash, and represent site priorities in regional and global IT initiatives.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
1 day ago

Senior Staff Site Reliability Operations

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 2 hours agoStatus: Live

Job Summary

Own day-to-day site reliability operations for NVIDIA’s Seattle, WA location, leading incidents, escalations, and service delivery quality. Serve as Tier 3 escalation across identity, messaging, endpoints, compute, and datacenter/lab hardware, driving root-cause analysis and permanent fixes. Lead standards and mentorship for site SRO engineers, manage compliance and patching, build automation with PowerShell/Python/Bash, and represent site priorities in regional and global IT initiatives.
Location: Seattle
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Sr. Manager level

Key Responsibilities

  • •Own day-to-day site operations, including incidents, requests, critical issues, support coverage, and accountability for queue health, SLA attainment, backlog, and service quality.
  • •Act as Tier 3 escalation for identity, messaging, compute, endpoint, and remediation work; drive root-cause analysis and permanent fixes.
  • •Own endpoint compliance, vulnerability remediation, patch management, hardening, audit readiness/evidence, and partner with InfoSec on incident response and privileged access.
  • •Lead site SRO engineers as technical lead by setting standards, reviewing work, directing blocking issues, mentoring, and maintaining the site knowledge base and runbooks.
  • •Build automation in PowerShell, Python, or Bash for diagnostics, remediation, health checks, and reporting; eliminate recurring drivers using ticket and reliability trend analysis.

Pay and Benefits

Salary: USD 184,000 - 264,500 annually
Equity and Bonus:Equity

Key Requirements

  • •12+ years in enterprise support engineering, infrastructure, or end user services, including 5+ years in a senior, lead, or escalation-tier role in a multi-site environment.
  • •Deep hands-on problem-solving across Active Directory, hybrid Entra ID, Exchange hybrid, Windows and Linux servers, virtualization, enterprise storage, and datacenter hardware.
  • •Endpoint management experience (Intune, Autopilot, MECM/SCCM, Jamf or equivalent), including Windows 11 and Microsoft 365 ecosystem with endpoint security and vulnerability remediation.
  • •Database operations support and networking fundamentals (DNS, DHCP, VLAN, wireless, firewall policy, and switch-level troubleshooting).
  • •Scripting/automation in Python, PowerShell, or Bash for support problems and ServiceNow (or similar ITSM); plus a Bachelor’s degree in CS/IS or equivalent experience.
Experience:12+ yearsEnterprise supportInfrastructureEnd user servicesMulti-site environment
Education:Bachelor's in Computer Science, Information Systems, or related field
Skills:Technical leadershipExecutive communicationRoot cause focusMentorshipDiagnostic rigor
Tech Stack:Active DirectoryHybrid Entra IDGPOKerberosLDAPSSOMFAExchange hybridSMTP relayWindowsLinuxMacOSVirtualizationStorageMicrosoft 365M365TeamsIntuneAutopilotPowerShell

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor