Principal Software Development Engineer - AI/ML

AMD
San Jose
Workplace: OnsiteFull timeUSD 210,000 - 360,000 annuallyFunction: Data Science & Machine LearningSkills: ["Problem-solving","Analytical skills","Communication","Ownership"]

Drive compiler and performance optimizations for state-of-the-art AI training and inference workloads on AMD GPU platforms. Collaborate across architecture, kernel, runtime, and ML framework teams to improve portability and deliver measurable gains through MLIR/LLVM and the ROCm software stack. Develop and maintain ONNX operators and microbenchmark test suites, and lead benchmarking across pre- and post-silicon environments to optimize compute efficiency, memory use, and communication overhead.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
AMD
AMD
1 day ago

Principal Software Development Engineer - AI/ML

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 8 hours agoStatus: Live

Job Summary

Drive compiler and performance optimizations for state-of-the-art AI training and inference workloads on AMD GPU platforms. Collaborate across architecture, kernel, runtime, and ML framework teams to improve portability and deliver measurable gains through MLIR/LLVM and the ROCm software stack. Develop and maintain ONNX operators and microbenchmark test suites, and lead benchmarking across pre- and post-silicon environments to optimize compute efficiency, memory use, and communication overhead.
Location: San Jose
Workplace: Onsite
Employment Type: Full time
Job Function: Data Science & Machine Learning
Seniority: Mid level

Key Responsibilities

  • •Develop and optimize compiler transformations and operator lowerings within MLIR, LLVM, and the ROCm software stack.
  • •Collaborate with compiler and runtime teams to design advanced optimization strategies for AI workloads on AMD GPUs.
  • •Implement, validate, and maintain ONNX operators and microbenchmark test suites.
  • •Lead performance optimization for AI training, fine-tuning, reinforcement learning (RL), and inference workloads on AMD GPUs.
  • •Drive performance characterization and benchmarking across pre-silicon and post-silicon environments, improving compute efficiency, memory utilization, and communication overhead.

Pay and Benefits

Salary: USD 210,000 - 360,000 annually

Key Requirements

  • •Strong compiler fundamentals and proven ability to improve performance for large-scale AI models.
  • •Strong proficiency in C++ and object-oriented software development.
  • •Hands-on GPU programming experience (e.g., HIP, CUDA, Triton or equivalent parallel frameworks).
  • •Experience with compiler technologies including MLIR and LLVM (or related compiler infrastructures).
  • •Bachelor’s or Master’s degree in Computer Science, Computer Engineering, Electrical Engineering, or equivalent.
Experience:AI trainingAI inferenceGPU programmingCompiler optimization
Education:
Skills:Problem-solvingAnalytical skillsCommunicationOwnership
Tech Stack:C++MLIRLLVMROCmONNX RuntimeONNXHIPCUDATritonLinuxGit

Eligibility

Work Authorization:Authorization required. Sponsorship not provided.

Company Brief

AMD
Designs and produces semiconductor products including CPUs, GPUs, and adaptive SoCs for consumer, enterprise, and embedded markets, competing across PCs, data centers, and gaming industries.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1969
Glassdoor
Glassdoor: 3.9
WebsiteLinkedIn