Solutions Architect - Physical AI TPM
NVIDIA
Beijing
Workplace: OnsiteFull timeFunction: Solutions Engineering & Sales EngineeringExperience: 10+ yearsEducation: mastersSkills: ["SGLang","VLLM","KV cache","FlexKV","Distributed training","HPC"]Lead architectural design for AI computing platforms focused on LLM inference and distributed training acceleration. Contribute to open-source inference frameworks (SGLang, vLLM), develop KV cache offloading (FlexKV), optimize performance for multi-node HPC environments, and prototype acceleration libraries to solve ML system bottlenecks. Work with global teams across research and engineering to deliver scalable, cutting-edge AI solutions.

