LLM Backend Engineer Graduate (Applied Machine Learning) - 2027 Start

ByteDance
San Jose
Workplace: OnsiteFull timeFunction: Data Science & Machine LearningEducation: bachelorsSkills: []

Build and develop core components of the Volcano Ark MaaS platform, including API capabilities for text dialogue, multimodal understanding, and multimodal generation. Help evolve cloud-native architectures (service mesh, load balancing, intelligent routing) and implement high-availability features like grayscale releases and resilience controls. Tune backend systems for long connections, streaming output, and low-latency multimodal inference, ensuring stability during large-scale traffic spikes.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
ByteDance
ByteDance
1 day ago

LLM Backend Engineer Graduate (Applied Machine Learning) - 2027 Start

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 7 hours agoStatus: Live

Job Summary

Build and develop core components of the Volcano Ark MaaS platform, including API capabilities for text dialogue, multimodal understanding, and multimodal generation. Help evolve cloud-native architectures (service mesh, load balancing, intelligent routing) and implement high-availability features like grayscale releases and resilience controls. Tune backend systems for long connections, streaming output, and low-latency multimodal inference, ensuring stability during large-scale traffic spikes.
Location: San Jose
Workplace: Onsite
Employment Type: Full time
Job Function: Data Science & Machine Learning
Seniority: Graduate level

Key Responsibilities

  • •Design and develop core components of the Volcano Ark MaaS platform, including API capabilities for text dialogue and multimodal understanding/generation.
  • •Participate in cloud-native architecture evolution (service mesh, load balancing, intelligent routing) and implement high-availability solutions such as grayscale release, traffic degradation, circuit breaking, and rate limiting.
  • •Perform system-level performance tuning for long connections, high throughput, streaming output, and low latency in multimodal large model inference.
  • •Ensure system stability for large-scale model invocation scenarios and resolve architectural bottlenecks from sudden traffic spikes.

Key Requirements

  • •Completing or recently completed a Bachelor's or Master's degree in Computer Science or a related discipline.
  • •Proficient in at least one programming language such as Golang, Java, C++, or Python, with practical software development experience.
  • •Understanding of backend engineering systems, including databases, computer networks, operating systems, and distributed system principles.
  • •Hands-on server-side engineering experience.
  • •Interest in large model inference pipelines with relevant background knowledge.
Education:Bachelor's in Computer Science
Tech Stack:GolangJavaC++PythonDatabasesComputer networksOperating systemsDistributed systemsKubernetesDockerIstioEnvoyService MeshLoad balancingRate limitingCircuit breakingGrayscale releaseTraffic degradation

Company Brief

ByteDance
Develops consumer internet and content platforms, including TikTok and other apps for short-form video, news, and entertainment. It also builds advertising, commerce, and creator tools that connect audiences, brands, and publishers across global markets.
Industry: Digital Media
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Established Company
Headquarters: Beijing, China
Founded: 2012
WebsiteLinkedIn