Senior GPU Inference Performance Engineer
AMD
Santa Clara
Workplace: HybridFull timeFunction: Solutions Engineering & Sales EngineeringEducation: bachelorsSkills: ["Evidence-driven","Rigorous","Communication","Written reporting","Collaboration"]Own end-to-end performance analysis for GPU-accelerated AI inference workloads. Profile and diagnose bottlenecks across GPU hardware and software runtimes, optimize inference engines and LLM serving frameworks, and run head-to-head benchmark comparisons. Analyze distributed inference networking and Kubernetes/GPU operator overhead, then automate trace collection and performance regression dashboards to present evidence-backed findings to product and executive stakeholders.

