AI/HPC Cluster Design Engineer

AMD
Austin
Workplace: OnsiteFull timeFunction: Design (Product/UX/UI/Visual)Education: bachelorsSkills: ["Collaboration","Problem-solving","Communication","Documentation","Strategic mindset"]

Design scalable AI/HPC clusters by selecting and reviewing compute, storage, networking, and power delivery components to meet customer and design requirements. Evaluate CPUs/GPUs, accelerators, interconnects, and memory configurations for performance and reliability across global deployments. Partner with cross-functional hardware, software, network, data center, and operations teams to deliver efficient, dependable infrastructure for AI and high-performance computing workloads.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
AMD
AMD
1 month ago

AI/HPC Cluster Design Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 16 hours agoStatus: Live

Job Summary

Design scalable AI/HPC clusters by selecting and reviewing compute, storage, networking, and power delivery components to meet customer and design requirements. Evaluate CPUs/GPUs, accelerators, interconnects, and memory configurations for performance and reliability across global deployments. Partner with cross-functional hardware, software, network, data center, and operations teams to deliver efficient, dependable infrastructure for AI and high-performance computing workloads.
Location: Austin
Workplace: Onsite
Employment Type: Full time
Job Function: Design (Product/UX/UI/Visual)

Key Responsibilities

  • •Design scalable AI/HPC clusters including compute, storage, and networking.
  • •Evaluate and select CPUs, GPUs, accelerators, interconnects, and memory configurations for optimal performance.
  • •Design network topologies and understand workload-specific network performance needs and trade-offs.
  • •Design and optimize cluster storage solutions, including understanding trade-offs for systems like Lustre and Ceph.
  • •Collaborate across hardware, software, network, data center, and operations teams to deliver reliable compute infrastructure.

Key Requirements

  • •Experienced systems engineer with strong background in HPC, AI systems, and cluster engineering.
  • •Deep technical knowledge of compute, power, and networking components for system-level design.
  • •Strong understanding of rack and cluster design.
  • •Knowledge of GPU/CPU architectures plus interconnect and networking technologies such as PCIe, UALink, InfiniBand, and Ethernet.
  • •Bachelor's or Master's degree in Electrical Engineering, Computer Engineering, Computer Science, or a related field.
Experience:HPCAI systemsData center engineering
Education:Bachelor's in Electrical Engineering, Computer Engineering, Computer Science or related field
Skills:CollaborationProblem-solvingCommunicationDocumentationStrategic mindset
Tech Stack:CPUsGPUsAcceleratorsInterconnectsMemory configurationsPCIeUALinkInfiniBandEthernetLustreCephAI/ML frameworks

Eligibility

Work Authorization:Authorization required. Sponsorship not provided.

Company Brief

AMD
Designs and produces semiconductor products including CPUs, GPUs, and adaptive SoCs for consumer, enterprise, and embedded markets, competing across PCs, data centers, and gaming industries.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1969
Glassdoor
Glassdoor: 3.9
WebsiteLinkedIn