Large Model Training Acceleration Engineer
ByteDance
San Jose
Workplace: OnsiteFull timeFunction: Education & TrainingExperience: 2+ yearsEducation: mastersSkills: []Build and optimize end-to-end large-model training and inference pipelines for ByteDance’s AI Platform team. You’ll improve training efficiency, speed, and scalability by designing distributed training strategies (data, model, and pipeline parallelism), and by benchmarking and profiling deep learning workloads to remove performance bottlenecks. Work on accelerating large-scale generative models to enhance performance and deployment at production scale.

