Principal ML Engineer - Large Scale Training Performance Optimization
San Jose
Workplace: OnsiteFull timeFunction: Data Science & Machine LearningEducation: mastersSkills: ["Communication","Problem-solving","Collaboration"]Lead efforts to train large AI models at scale on AMD GPUs, optimizing distributed training pipelines and end-to-end performance. Collaborate across teams to push the AMD AI platform forward, contribute to open source, and stay at the forefront of training algorithms and techniques for large-scale models.
Loading
Loading job details...
Preparing the role view and application actions.

