Research Scientist, Applied GAI-Vision

ByteDance
San Jose
Workplace: OnsiteFull timeFunction: Data Science & Machine LearningExperience: 2+ yearsSkills: ["Algorithms","Programming","Research","Technology transfer"]

Apply cutting-edge research in computer vision and generative AI to build capabilities for content creation and multimodal understanding. You’ll conduct R&D across generative models (diffusion/GAN), image/video synthesis and editing, human modeling, and facial analysis, then transfer technologies into ByteDance products. Collaborate with an applied vision research team exploring AI-first products that enable users to make and share creative content.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
ByteDance
ByteDance
1 month ago

Research Scientist, Applied GAI-Vision

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 30 days agoStatus: Live
Reposted: similar role first listed 1 month ago

Job Summary

Apply cutting-edge research in computer vision and generative AI to build capabilities for content creation and multimodal understanding. You’ll conduct R&D across generative models (diffusion/GAN), image/video synthesis and editing, human modeling, and facial analysis, then transfer technologies into ByteDance products. Collaborate with an applied vision research team exploring AI-first products that enable users to make and share creative content.
Location: San Jose
Workplace: Onsite
Employment Type: Full time
Job Function: Data Science & Machine Learning

Key Responsibilities

  • •Conduct cutting-edge research and development in computer vision, generative AI, and related fields.
  • •Transfer advanced research technologies into ByteDance products.
  • •Explore new AI-first product ideas where artificial intelligence is at the core.
  • •Work across research groups spanning generative models for content creation, image generation, video synthesis, intelligent image/video editing, and virtual humans.

Key Requirements

  • •2+ years of research and practical experience in computer vision areas such as generative models (diffusion models, GAN), image/video synthesis or understanding, multi-modality, or generative human and facial analysis.
  • •Highly competent in algorithms and programming.
  • •Strong coding skills in C/C++ and Python.
  • •Experience with one or more areas including image/video editing, vision and language, neural motion synthesis, pose estimation, hand pose estimation, or human shape estimation.
  • •Publications in venues such as CVPR, ICCV, ECCP, SIGGRAPH, or NeurIPS (preferred).
Experience:2+ years
Skills:AlgorithmsProgrammingResearchTechnology transfer
Tech Stack:C/C++PythonComputer visionGenerative AIDiffusion modelsGANMulti-modalityImage and video synthesisImage and video understandingImage and video editingPose estimationHand pose estimationHuman shape estimation

Company Brief

ByteDance
Develops consumer internet and content platforms, including TikTok and other apps for short-form video, news, and entertainment. It also builds advertising, commerce, and creator tools that connect audiences, brands, and publishers across global markets.
Industry: Digital Media
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Established Company
Headquarters: Beijing, China
Founded: 2012
WebsiteLinkedIn