Senior Infrastructure Automation Engineer, Compute Platform - EDA Infrastructure

NVIDIA
North Carolina, Austin
Workplace: HybridFull timeUSD 184,000 - 356,500 annuallyFunction: QA, Test & Release EngineeringEducation: bachelorsSkills: []

Own a config-as-code automation foundation for an EDA compute farm, migrating from partially manual scheduler configurations to consistent, policy-driven deployment. Design the configuration schema for LSF cell deployment, build a merge-to-production pipeline with staged rollouts and rollback, and eliminate configuration drift across a federated estate. Stand up regression testing to support scheduled LSF upgrades, and encode scheduler expertise into templates that live in the repository.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
1 week ago

Senior Infrastructure Automation Engineer, Compute Platform - EDA Infrastructure

âś“ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 1 hour agoStatus: Live

Job Summary

Own a config-as-code automation foundation for an EDA compute farm, migrating from partially manual scheduler configurations to consistent, policy-driven deployment. Design the configuration schema for LSF cell deployment, build a merge-to-production pipeline with staged rollouts and rollback, and eliminate configuration drift across a federated estate. Stand up regression testing to support scheduled LSF upgrades, and encode scheduler expertise into templates that live in the repository.
Location: North Carolina, Austin
Workplace: Hybrid
Employment Type: Full time
Job Function: QA, Test & Release Engineering
Seniority: Mid level

Key Responsibilities

  • •Design and own the configuration schema for LSF cell deployment so policy changes are written once and applied identically everywhere.
  • •Build the deployment pipeline to move scheduler configuration from merge to production across a federated estate, including staged rollout and rollback.
  • •Eliminate configuration drift across scheduler cells and create tooling to keep drift eliminated.
  • •Stand up a regression suite enabling scheduled LSF upgrades rather than ad-hoc upgrades.
  • •Collaborate with an LSF internals engineer to encode scheduler knowledge into repository templates and policy.

Pay and Benefits

Salary: USD 184,000 - 356,500 annually
Equity and Bonus:Equity

Key Requirements

  • •BS or MS in Computer Science or equivalent experience.
  • •6+ years in infrastructure engineering with strong config management depth (Ansible, Salt, Puppet, Chef, or comparable).
  • •Real experience with GitOps at scale, including review workflow, environment promotion, drift detection, and safe rollback.
  • •Proficiency in Go, Python, and shell, and comfort building automation tooling.
  • •Experience automating stateful, long-lived infrastructure that can’t simply be destroyed and recreated.
Experience:Infrastructure engineeringConfig managementGitOpsBatch schedulingEDA
Education:Bachelor's in Computer Science
Tech Stack:LSFSlurmAnsibleSaltPuppetChefGitOpsGoPythonShellCI/CDConfig managementDrift detectionStaged rolloutRollback

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor