Senior Datacenter Platform/Debug Engineer

AMD
United States
Workplace: OnsiteFull timeUSD 98,560 - 168,960 annuallyFunction: IT Operations (Systems/Network Admin)Education: mastersSkills: ["Troubleshooting","Analytical thinking","Collaboration","Communication","Documentation"]

Join AMD’s Datacenter Infrastructure team to deploy, operate, and debug next-generation HPC and AI server and rack-scale platforms. You’ll serve as a technical leader in the lab environment—deploying complex systems, executing hardware installs, and performing root-cause analysis across hardware, firmware, software, OS, and networking. You’ll also build scripts/tools for efficiency, document investigations, and collaborate across hardware, software, validation, and infrastructure teams, with occasional travel to AMD datacenter sites.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
AMD
AMD
6 days ago

Senior Datacenter Platform/Debug Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 2 hours agoStatus: Live

Job Summary

Join AMD’s Datacenter Infrastructure team to deploy, operate, and debug next-generation HPC and AI server and rack-scale platforms. You’ll serve as a technical leader in the lab environment—deploying complex systems, executing hardware installs, and performing root-cause analysis across hardware, firmware, software, OS, and networking. You’ll also build scripts/tools for efficiency, document investigations, and collaborate across hardware, software, validation, and infrastructure teams, with occasional travel to AMD datacenter sites.
Location: United States
Workplace: Onsite
Employment Type: Full time
Job Function: IT Operations (Systems/Network Admin)
Seniority: Mid level

Key Responsibilities

  • •Deploy, configure, and maintain server, rack, and datacenter infrastructure platforms.
  • •Execute hardware installation activities supporting lab and datacenter environments.
  • •Troubleshoot and debug hardware, firmware, software, operating system, and networking issues.
  • •Perform root cause analysis of complex system failures and drive resolution efforts.
  • •Develop and maintain scripts and tools to improve operational efficiency and automation.

Pay and Benefits

Salary: USD 98,560 - 168,960 annually

Key Requirements

  • •Experience supporting server, datacenter, HPC, or AI infrastructure environments.
  • •Knowledge of computer hardware architecture (CPUs, GPUs, memory subsystems, networking, I/O devices).
  • •Experience debugging hardware, firmware, operating systems, and software applications.
  • •Linux system administration experience.
  • •Bachelor’s or master’s degree in Computer Engineering, Electrical Engineering, Computer Science, or a related technical discipline (preferred).
Experience:HPCAI infrastructureDatacenterServer hardware
Education:Master's in Computer Engineering, Electrical Engineering, Computer Science, or a related technical discipline
Skills:TroubleshootingAnalytical thinkingCollaborationCommunicationDocumentation
Languages:English
Tech Stack:PythonBashCC++LinuxPCIeRedfishIPMII2CSPINVLinkXGMIInfinity Fabric

Eligibility

Work Authorization:Authorization required. Sponsorship not provided.

Company Brief

AMD
Designs and produces semiconductor products including CPUs, GPUs, and adaptive SoCs for consumer, enterprise, and embedded markets, competing across PCs, data centers, and gaming industries.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1969
Glassdoor
Glassdoor: 3.9
WebsiteLinkedIn