Software Engineer, LLM Inference
Shanghai
Workplace: OnsiteFull timeFunction: Software EngineeringExperience: 4+ yearsEducation: mastersSkills: ["Communication","Customer communication","Collaboration","Proactivity","Debugging"]Develop and scale robust LLM inference software across multiple platforms, focusing on functionality, performance, and tuning. Conduct performance analysis and optimization, and stay current with AI/LLM research to improve model inferencing. Work on TensorRT and TensorRT Edge LLM updates, and collaborate with software, research, and product teams to guide machine-learning inferencing direction.
Loading
Loading job details...
Preparing the role view and application actions.

