Software Engineer Graduate (Data-Speech-Product RD-Engineering-US) - 2027 Start

ByteDance
San Jose
Workplace: OnsiteFull timeFunction: Software EngineeringEducation: mastersSkills: ["Analytical skills","Teamwork"]

Join the Speech team to design and implement AI model optimization techniques that improve speed, efficiency, and scalability for production. You’ll build benchmarking frameworks, optimize training and inference pipelines on GPUs and distributed systems, and collaborate with ML researchers to transition optimized models into product environments. The role focuses on areas like quantization, pruning, knowledge distillation, and efficient architectures while keeping up with advances in compilers and systems.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
ByteDance
ByteDance
2 hours ago

Software Engineer Graduate (Data-Speech-Product RD-Engineering-US) - 2027 Start

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 2 hours agoStatus: Live

Job Summary

Join the Speech team to design and implement AI model optimization techniques that improve speed, efficiency, and scalability for production. You’ll build benchmarking frameworks, optimize training and inference pipelines on GPUs and distributed systems, and collaborate with ML researchers to transition optimized models into product environments. The role focuses on areas like quantization, pruning, knowledge distillation, and efficient architectures while keeping up with advances in compilers and systems.
Location: San Jose
Workplace: Onsite
Employment Type: Full time
Job Function: Software Engineering
Seniority: Graduate level

Key Responsibilities

  • •Develop and implement algorithms for model optimization, including quantization, pruning, knowledge distillation, and efficient architectures.
  • •Build and maintain performance benchmarking frameworks for large-scale training and inference.
  • •Optimize training and inference pipelines on GPUs and across distributed systems.
  • •Collaborate with ML researchers to transition optimized models into production.
  • •Stay current with the latest research in model efficiency, compilers, and systems.

Key Requirements

  • •Currently completing or recently completed a Bachelor's or Master's degree in computer engineering or a related discipline.
  • •Strong coding skills in Python and C++.
  • •Ability to optimize training and inference workflows in high-performance environments.
  • •Understanding of computer architecture, parallel computing, and GPU acceleration.
  • •Analytical skills and ability to work in a fast-paced team environment.
Experience:Deep learningDistributed training
Education:Master's in computer engineering
Skills:Analytical skillsTeamwork
Tech Stack:PythonC++CUDATritonTVMXLATensorRTGPUsDistributed systemsQuantizationPruningKnowledge distillation

Eligibility

Work Authorization:Authorization required. Sponsorship not provided.

Company Brief

ByteDance
Develops consumer internet and content platforms, including TikTok and other apps for short-form video, news, and entertainment. It also builds advertising, commerce, and creator tools that connect audiences, brands, and publishers across global markets.
Industry: Digital Media
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Established Company
Headquarters: Beijing, China
Founded: 2012
WebsiteLinkedIn