Research Scientist - Technologies of Data Management, LLM and AI Agents - Global Frontier Tech Recruitment Program - 2027 Start (PhD)

ByteDance
Seattle
Workplace: OnsiteFull timeFunction: Data Science & Machine LearningEducation: phdSkills: ["Learning","Self-motivation","Teamwork","Communication","Problem analysis"]

Conduct end-to-end research on AI-native infrastructure that supports large-scale LLMs and AI agents. Work on data management for LLM/agent workloads, including multimodal query processing (vector/full-text/SQL) and cost- and latency-efficient semantic operators. Build intelligent infrastructure optimization from agent workflows, covering network observability, storage acceleration, power scheduling, and low-latency vector retrieval to improve utilization and reduce costs.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
ByteDance
ByteDance
1 month ago

Research Scientist - Technologies of Data Management, LLM and AI Agents - Global Frontier Tech Recruitment Program - 2027 Start (PhD)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 30 days agoStatus: Live
Reposted: similar role first listed 1 month ago

Job Summary

Conduct end-to-end research on AI-native infrastructure that supports large-scale LLMs and AI agents. Work on data management for LLM/agent workloads, including multimodal query processing (vector/full-text/SQL) and cost- and latency-efficient semantic operators. Build intelligent infrastructure optimization from agent workflows, covering network observability, storage acceleration, power scheduling, and low-latency vector retrieval to improve utilization and reduce costs.
Location: Seattle
Workplace: Onsite
Employment Type: Full time
Job Function: Data Science & Machine Learning
Seniority: Graduate level

Key Responsibilities

  • •Research and develop data management technologies for LLMs and AI agents, including multimodal query processing over vector/full-text/SQL data.
  • •Design cost-effective, low-latency semantic operators and scalable data processing for typed datasets.
  • •Explore infrastructure auto-optimization driven by AI agent workflows, enabling intelligent optimization through “AI for Infrastructure.”
  • •Investigate AI infrastructure stack topics including network observability, storage systems, and GPU/CPU/MEM scheduling.
  • •Build and optimize vector retrieval components, including a cloud-native distributed vector index engine for low-latency, low-cost retrieval.

Key Requirements

  • •Completing or recently completed a PhD in Software Development, Computer Science, Computer Engineering, or a related technical discipline.
  • •Strong coding ability in at least one mainstream programming language (e.g., C/C++, Python, Go), with data structures and algorithms fundamentals.
  • •Ability to independently design and develop complex systems and produce design documents and deliverable demo systems.
  • •Familiarity with state-of-the-art data management technologies.
  • •Research experience in LLMs and infrastructure (preferred), with strong learning ability and problem analysis skills.
Education:PhD / Doctorate
Skills:LearningSelf-motivationTeamworkCommunicationProblem analysis
Tech Stack:C/C++PythonGoSQLTextToSQLNL-to-SQLLLMsAI AgentsVector retrievalTime-series databasesServerlessDPUGPUCPUMEM

Company Brief

ByteDance
Develops consumer internet and content platforms, including TikTok and other apps for short-form video, news, and entertainment. It also builds advertising, commerce, and creator tools that connect audiences, brands, and publishers across global markets.
Industry: Digital Media
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Established Company
Headquarters: Beijing, China
Founded: 2012
WebsiteLinkedIn