AI Framework Engineer

AMD
Shanghai
Full timeCNY 581,350 - 830,500 annuallyFunction: Data Science & Machine LearningExperience: 5+ yearsEducation: mastersSkills: ["Problem-solving","Proactive approach","Collaboration","Mentorship","Engineering best practices"]

Develop and optimize deep learning frameworks and GPU kernels for AMD GPUs. Enhance performance for training and inference across multi-GPU and multi-node systems by improving frameworks such as TensorFlow and PyTorch and integrating optimizations into open-source repositories. Collaborate with internal GPU library teams and open-source maintainers, leverage advanced compiler technologies, and improve the end-to-end deep learning pipeline, including graph compilers. Provide mentorship through code reviews and technical guidance.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
AMD
AMD
13 hours ago

AI Framework Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 10 hours agoStatus: Live

Job Summary

Develop and optimize deep learning frameworks and GPU kernels for AMD GPUs. Enhance performance for training and inference across multi-GPU and multi-node systems by improving frameworks such as TensorFlow and PyTorch and integrating optimizations into open-source repositories. Collaborate with internal GPU library teams and open-source maintainers, leverage advanced compiler technologies, and improve the end-to-end deep learning pipeline, including graph compilers. Provide mentorship through code reviews and technical guidance.
Location: Shanghai
Employment Type: Full time
Job Function: Data Science & Machine Learning
Seniority: Mid level

Key Responsibilities

  • •Optimize deep learning frameworks such as TensorFlow and PyTorch for AMD GPUs in open-source repositories.
  • •Create and optimize GPU kernels to maximize performance for specific AI operations.
  • •Design and optimize deep learning models for AMD GPU performance.
  • •Collaborate with internal GPU library teams to analyze and improve training and inference performance.
  • •Work with open-source maintainers to align code changes with requirements and integrate upstream. Make use of advanced compiler technologies and improve the deep learning pipeline, including graph compilers.

Pay and Benefits

Salary: CNY 581,350 - 830,500 annually

Key Requirements

  • •5+ years of professional experience in technical software development with a focus on GPU optimization, performance engineering, and framework development.
  • •Master’s degree in Computer Science, Computer Engineering, Electrical Engineering, or a related field.
  • •Strong expertise in C++ development within Linux environments.
  • •Expert skills in Python and C++, including debugging, performance tuning, and test design for maintainable software.
  • •Experience designing and optimizing GPU kernels for deep learning on AMD GPUs using HIP, CUDA, and assembly (ASM), with knowledge of AMD architectures (GCN, RDNA).
Experience:5+ yearsAIDeep learningGPU optimizationPerformance engineeringFramework development
Education:Master's in Computer Science
Skills:Problem-solvingProactive approachCollaborationMentorshipEngineering best practices
Languages:English
Tech Stack:C++LinuxPythonTensorFlowPyTorchAMD GPUsGPU kernelsHIPCUDAAssembly (ASM)GCNRDNACompute Kernel (CK)CUTLASSTritonMulti-GPUMulti-nodeCompiler technologiesGraph compilersLLVM

Company Brief

AMD
Designs and produces semiconductor products including CPUs, GPUs, and adaptive SoCs for consumer, enterprise, and embedded markets, competing across PCs, data centers, and gaming industries.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1969
Glassdoor
Glassdoor: 3.9
WebsiteLinkedIn