Distinguished Engineer, Production Engineering, Cluster Management
Santa Clara, Oregon
Workplace: HybridFull timeUSD 320,000 - 488,750 annuallyFunction: Manufacturing & Production OperationsExperience: 18+ yearsSkills: ["Architectural judgment","Technical leadership","Cross-team influence","Operational reliability thinking","Automation mindset"]Lead production engineering for DGX Cloud GPU capacity as a hands-on technical authority. Define long-range strategy and architectural direction for cluster lifecycle, runtime delivery, restoration, release readiness, and steady-state operability across on-prem, hyperscalers, and NVIDIA Cloud Partner environments. Build durable workflows, APIs, and automation for Kubernetes-based service management and readiness gates, and drive cross-team improvements in reliability, operability, performance, and release safety.
Loading
Loading job details...
Preparing the role view and application actions.

