PMTS Systems Design Engineer

AMD
Santa Clara
Workplace: RemoteFull timeUSD 200,900 - 344,400 annuallyFunction: Design (Product/UX/UI/Visual)Education: mastersSkills: ["Debugging","Root cause analysis","Benchmarking","Performance tuning"]

Design, develop, and test operating systems and supporting software tools for semiconductor and GPU-to-GPU networking used in AI training and inferencing. Collaborate with hardware and software teams to improve GPU cluster performance across RDMA throughput, latency, and collective communications. Lead debugging and root-cause analysis, build benchmarking strategies, and enhance Linux networking debug capabilities using scripting and system-level performance analysis tools.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
AMD
AMD
3 days ago

PMTS Systems Design Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 3 hours agoStatus: Live

Job Summary

Design, develop, and test operating systems and supporting software tools for semiconductor and GPU-to-GPU networking used in AI training and inferencing. Collaborate with hardware and software teams to improve GPU cluster performance across RDMA throughput, latency, and collective communications. Lead debugging and root-cause analysis, build benchmarking strategies, and enhance Linux networking debug capabilities using scripting and system-level performance analysis tools.
Location: Santa Clara
Workplace: Remote
Employment Type: Full time
Job Function: Design (Product/UX/UI/Visual)
Seniority: Mid level

Key Responsibilities

  • •Research, design, develop, and/or test operating systems for semiconductor operations and associated software systems/tools.
  • •Integrate software solutions enabling Ethernet to support GPU-to-GPU communication for AI inferencing and training.
  • •Collaborate with hardware and software teams to enhance GPU cluster performance (e.g., RDMA throughput, latency, collective communications).
  • •Provide root cause analysis guidance and lead debug efforts to identify root causes and resolution.
  • •Develop and execute benchmarking strategies to assess baseline performance and identify bottlenecks and improvements for GPU cluster environments.

Pay and Benefits

Salary: USD 200,900 - 344,400 annually

Key Requirements

  • •Master’s degree in Computer Science, Computer Engineering, Electrical Engineering, Telecommunications, or a related field.
  • •5 years of experience in the job offered or closely related engineering role.
  • •Performance optimization of GPU clusters and understanding of GPU architectures, parallel computing concepts, and network protocols.
  • •Experience with scripting languages, including Python and Bash, for automation and performance analysis.
  • •Experience with system-level performance analysis tools and methodologies, software debugging, Linux kernel networking, and RDMA network configuration/troubleshooting/performance tuning.
Experience:HPCMachine learningGPULinuxNetworking
Education:Master's in Computer Science, Computer Engineering, Electrical Engineering, Telecommunications, or related field
Skills:DebuggingRoot cause analysisBenchmarkingPerformance tuning
Languages:En-us
Tech Stack:PythonBashLinux kernel networkingRDMAMachine learningHPCGPU clusterNetwork protocolsEthernetCloud

Company Brief

AMD
Designs and produces semiconductor products including CPUs, GPUs, and adaptive SoCs for consumer, enterprise, and embedded markets, competing across PCs, data centers, and gaming industries.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1969
Glassdoor
Glassdoor: 3.9
WebsiteLinkedIn