Software Engineer - AI Compute Infrastructure
ByteDance
San Jose
Workplace: OnsiteFull timeFunction: Software EngineeringExperience: 2+ yearsEducation: bachelorsSkills: ["Communication","Collaboration","Performance optimization","Open-source innovation"]Build and operate cloud-native, GPU-optimized inference infrastructure for large-scale LLM workloads. You’ll design container-based cluster management and orchestration systems, architect secure and cost-efficient GPU/AI accelerator platforms, and collaborate on inference solutions using vLLM, SGLang, and TensorRT-LLM. Work in a hyper-scale environment, integrate advances from Kubernetes/Ray and ML systems research, and contribute production-ready, maintainable code.

