Large Model Algorithm Researcher (Multimodal & Code AI)- Soaring Star Talent Program

ByteDance
Singapore
Workplace: OnsiteFull timeFunction: Research & Scientific (R&D)Education: phdSkills: ["Problem analysis","Problem-solving","Communication","Team spirit"]

Research and develop multimodal foundation large models and Code AI capabilities for TikTok business scenarios. Work on challenges such as improving multimodal perception encoders with adaptive frame rates and integrating modalities like audio and user behavior. Advance model perception and reasoning by fusing multimodal inputs with thinking capabilities, contributing to next-generation multilingual and video understanding algorithms.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
ByteDance
ByteDance
1 month ago

Large Model Algorithm Researcher (Multimodal & Code AI)- Soaring Star Talent Program

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 30 days agoStatus: Live
Reposted: similar role first listed 1 month ago

Job Summary

Research and develop multimodal foundation large models and Code AI capabilities for TikTok business scenarios. Work on challenges such as improving multimodal perception encoders with adaptive frame rates and integrating modalities like audio and user behavior. Advance model perception and reasoning by fusing multimodal inputs with thinking capabilities, contributing to next-generation multilingual and video understanding algorithms.
Location: Singapore
Workplace: Onsite
Employment Type: Full time
Job Function: Research & Scientific (R&D)

Key Responsibilities

  • •Enhance the multimodal perception encoder beyond fixed frame-rate approaches by exploring adaptive frame rates.
  • •Integrate additional modalities such as audio and user behavior into multimodal perception.
  • •Fuse multimodal perception with model thinking to strengthen comprehensive perception and cognitive abilities.
  • •Develop and improve multimodal foundation large models for multilingual and massive video content understanding.
  • •Apply Code AI research techniques to improve code understanding and reasoning for better program performance and R&D efficiency.

Key Requirements

  • •PhD degree, with priority for published papers in machine learning (ML), computer vision (CV), and natural language processing (NLP).
  • •Strong programming skills with data structures and algorithms, proficient in C/C++ or Python.
  • •Research experience in machine learning, especially large-scale language models (LLMs) and generative AI.
  • •Awards/competition track record (e.g., ACM/ICPC, NOI/IOI, Top Coder, Kaggle) is preferred.
  • •Passionate about technology with strong problem analysis/solving, communication skills, and team spirit.
Education:PhD / Doctorate
Skills:Problem analysisProblem-solvingCommunicationTeam spirit
Tech Stack:CC++PythonMachine learningComputer visionNatural language processingLLMsMultimodal large modelsMultimodal perceptionGenerative AI

Company Brief

ByteDance
Develops consumer internet and content platforms, including TikTok and other apps for short-form video, news, and entertainment. It also builds advertising, commerce, and creator tools that connect audiences, brands, and publishers across global markets.
Industry: Digital Media
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Established Company
Headquarters: Beijing, China
Founded: 2012
WebsiteLinkedIn