Deep Learning Performance Architect
NVIDIA
Shanghai, Beijing
Workplace: OnsiteFull timeFunction: Data Science & Machine LearningExperience: 2+ yearsEducation: phdSkills: ["Communication","Cross-functional collaboration","Autonomy"]Build and optimize GPU-accelerated deep learning inference software, including highly optimized inference kernels and performance tuning. You’ll analyze and profile workloads, apply architectural knowledge of CPUs and GPUs, and collaborate with cross-functional teams across automotive, image understanding, and speech understanding. Work with the deep learning community to help implement and release the latest algorithms in TensorRT. Occasionally travel for technical consultation and training.

