Triton Compiler/GPU Kernel Performance Engineer
Shanghai
Workplace: OnsiteFull timeFunction: Solutions Engineering & Sales EngineeringSkills: ["Communication","Leadership","Problem-solving"]Kernel Performance Architect responsible for defining, analyzing, and optimizing performance across the full stack—from GPU microarchitecture and compiler behavior to runtime systems and deep learning frameworks—for AI workloads on AMD GPUs. Lead cross-team efforts on cross-architecture optimization, performance modeling, and portability, while guiding implementation engineers and communicating tradeoffs.
Loading
Loading job details...
Preparing the role view and application actions.

