Principal Engineer, Efficient GenAI
San Jose, Seattle, Austin
Workplace: HybridFull timeUSD 210,000 - 360,000 annuallyFunction: Communications, PR & CommunityEducation: mastersSkills: ["Written communication","Verbal communication","Presentation skills","Coordination","Publishing externally"]Build and scale efficient Generative AI training and inference for large foundation models as part of AMD’s AI Models and Applications team. Drive innovations across transformer architectures, distributed training parallelism, and inference optimizations like speculative decoding and KV-caching. Co-optimize end-to-end performance with software and hardware teams, integrate AMD-optimized model libraries using open-source frameworks, and publish results at external conferences.
Loading
Loading job details...
Preparing the role view and application actions.

