AI/LLM Network Research Intern (High Speed Network) - 2027 Start (PhD)

ByteDance
San Jose
Workplace: OnsiteInternshipFunction: Research & Scientific (R&D)Education: phdSkills: []

Design and deploy high-performance transport protocols and congestion control algorithms to enable AI/LLM applications. Conduct R&D on AI communication framework, network protocol stacks, and host-network-application co-design to improve scalability, reliability, and performance of AI/LLM networks. Stay current with academia and industry advances and communicate results through academic-style presentations and papers, while gaining hands-on experience in hyperscale data-center networking.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
ByteDance
ByteDance
2 hours ago

AI/LLM Network Research Intern (High Speed Network) - 2027 Start (PhD)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 2 hours agoStatus: Live

Job Summary

Design and deploy high-performance transport protocols and congestion control algorithms to enable AI/LLM applications. Conduct R&D on AI communication framework, network protocol stacks, and host-network-application co-design to improve scalability, reliability, and performance of AI/LLM networks. Stay current with academia and industry advances and communicate results through academic-style presentations and papers, while gaining hands-on experience in hyperscale data-center networking.
Location: San Jose
Workplace: Onsite
Employment Type: Internship
Job Function: Research & Scientific (R&D)
Seniority: Intern level

Key Responsibilities

  • •Design, optimize, implement, and deploy high-performance transport protocols to support AI/LLM applications.
  • •Design, optimize, implement, and deploy congestion control algorithms to support AI/LLM applications.
  • •Research and develop high-performance AI communication frameworks and network protocol stacks, including host-network-application co-design for improved AI/LLM network scalability, reliability, and performance.
  • •Follow the latest technologies from academia and industry, identify innovative system components, and present findings in academic papers.

Key Requirements

  • •Currently pursuing a PhD in computer networking or a related technical discipline.
  • •Experience with network protocols such as TCP and RoCEv2, including network programming using Socket and verbs APIs, or familiarity with data center congestion control algorithms and their tradeoffs.
  • •Knowledge of scale-up protocols such as PCIe, NVLink, and UALink and how they differ from scale-out network protocols.
  • •Experience in high-speed network systems, including RDMA, congestion control, and AI network optimization.
  • •Proficiency in one or more programming languages such as C/C++, Python, or Go, and some knowledge of GPU architecture.
Education:PhD / Doctorate in Computer networking or a related technical discipline
Tech Stack:AI/LLMNetwork protocolsTCPRoCEv2SocketVerbs APIsCongestion control algorithmsPCIeNVLinkUALinkHigh-speed networksRDMAGPU architectureC/C++PythonGoSDNNetwork virtualizationNCCLMPI

Company Brief

ByteDance
Develops consumer internet and content platforms, including TikTok and other apps for short-form video, news, and entertainment. It also builds advertising, commerce, and creator tools that connect audiences, brands, and publishers across global markets.
Industry: Digital Media
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Established Company
Headquarters: Beijing, China
Founded: 2012
WebsiteLinkedIn