Research Scientist Intern - Multimodal Sensing & On-Device Perception - Global Frontier Tech Recruitment Program - 2027 Start (PHD)

ByteDance
San Jose
Workplace: OnsiteInternshipFunction: Data Science & Machine LearningEducation: phdSkills: ["Research","Model training","Self-motivation","Curiosity","Systems thinking"]

Work on an eye-tracking system architecture spanning sensing and computing, aiming for high coverage, performance, and low power. Design and prototype novel sensor/imaging architectures, build end-to-end imaging pipelines (optics/sensor → ISP → perception models), and co-optimize machine vision models with hardware constraints like power, bandwidth, and latency. Leverage VLM/LLM and world models to guide what sensing should preserve or transform.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
ByteDance
ByteDance
1 month ago

Research Scientist Intern - Multimodal Sensing & On-Device Perception - Global Frontier Tech Recruitment Program - 2027 Start (PHD)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 30 days agoStatus: Live
Reposted: similar role first listed 1 month ago

Job Summary

Work on an eye-tracking system architecture spanning sensing and computing, aiming for high coverage, performance, and low power. Design and prototype novel sensor/imaging architectures, build end-to-end imaging pipelines (optics/sensor → ISP → perception models), and co-optimize machine vision models with hardware constraints like power, bandwidth, and latency. Leverage VLM/LLM and world models to guide what sensing should preserve or transform.
Location: San Jose
Workplace: Onsite
Employment Type: Internship
Job Function: Data Science & Machine Learning
Seniority: Intern level

Key Responsibilities

  • •Design and prototype novel sensor or imaging architectures that move computation closer to the sensing front-end (e.g., near-sensor processing, event-driven capture, learned compression at the pixel level).
  • •Build and characterize imaging pipelines end-to-end: from optical/sensor physics through ISP to downstream perception models, identifying where bits are wasted and where intelligence should be injected.
  • •Leverage understanding of VLM/LLM and world models to determine what information the sensing front-end must preserve, discard, or transform.
  • •Develop or adapt machine vision models co-optimized with hardware constraints including power, bandwidth, and latency.

Key Requirements

  • •Currently pursuing a PhD in Computer Science, Electrical Engineering, Optical Engineering, Applied Mathematics, Physics, or a related technical field.
  • •Strong research background in computer vision and machine learning, with hands-on model training experience.
  • •Experience with at least one of: sequence modeling, language modeling, efficient neural network design, or signal processing.
  • •Proven track record of high-impact research, demonstrated by publications in CVPR, ICCV, ECCV, NeurIPS, ICLR, SIGGRAPH, or similar.
  • •Hands-on experience with hardware prototyping involving cameras, structured light, or other active sensing systems.
Experience:Computer visionMachine learningOn-device ML
Education:PhD / Doctorate
Skills:ResearchModel trainingSelf-motivationCuriositySystems thinking
Tech Stack:Computer visionMachine learningSequence modelingLanguage modelingEfficient neural networksSignal processingVLMLLMWorld modelsSensorsImage sensorsRefractive/diffractive opticsISPEvent-driven captureLearned compressionEmbedded ML deploymentHardware/software co-design

Company Brief

ByteDance
Develops consumer internet and content platforms, including TikTok and other apps for short-form video, news, and entertainment. It also builds advertising, commerce, and creator tools that connect audiences, brands, and publishers across global markets.
Industry: Digital Media
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Established Company
Headquarters: Beijing, China
Founded: 2012
WebsiteLinkedIn