Research Engineer - Data Infrastructure

ElevenLabs
London, New York, San Francisco, Warsaw, Bulgaria
Workplace: RemoteFull timeFunction: Research & Scientific (R&D)Skills: ["Problem-solving","Autonomous evaluation"]

Own the data infrastructure that powers ElevenLabs’ frontier AI models, from building large-scale pipelines to collecting, processing, filtering, and transforming training datasets. Develop and train models used in those pipelines, and design curation strategies (deduplication, quality scoring, labeling, augmentation) that measurably improve model outcomes. Create tooling and reliable infrastructure so researchers can explore and train on massive datasets quickly, using distributed systems at scale.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
ElevenLabs
ElevenLabs
4 days ago

Research Engineer - Data Infrastructure

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 7 hours agoStatus: Live

Job Summary

Own the data infrastructure that powers ElevenLabs’ frontier AI models, from building large-scale pipelines to collecting, processing, filtering, and transforming training datasets. Develop and train models used in those pipelines, and design curation strategies (deduplication, quality scoring, labeling, augmentation) that measurably improve model outcomes. Create tooling and reliable infrastructure so researchers can explore and train on massive datasets quickly, using distributed systems at scale.
Location: London, New York, San Francisco, Warsaw, Bulgaria
Workplace: Remote
Employment Type: Full time
Job Function: Research & Scientific (R&D)

Key Responsibilities

  • •Build large-scale data pipelines for collecting, processing, filtering, and transforming datasets used to train state-of-the-art models.
  • •Train models used within data processing pipelines (e.g., classifiers, quality filters, labeling models).
  • •Design data curation strategies such as deduplication, quality scoring, labeling, and augmentation to improve model performance.
  • •Create tooling and infrastructure that enables researchers to explore and train on massive datasets quickly and reliably.

Pay and Benefits

Perks:Learning BudgetCo-working StipendAnnual Offsite

Key Requirements

  • •Experience building data-intensive systems, ideally supporting machine learning training pipelines.
  • •Strong engineering skills in distributed data processing at scale (e.g., Kubernetes or custom pipelines).
  • •Ability to evaluate how data quality, composition, and curation affect model outcomes and build measurement tooling.
  • •Capability to solve hard technical problems and demonstrate it via past projects, designs, or GitHub contributions.
Experience:Machine learningData infrastructureAI
Skills:Problem-solvingAutonomous evaluation
Tech Stack:KubernetesGitHub

Company Brief

ElevenLabs
Develops advanced AI audio models and tools for realistic text-to-speech, voice cloning, dubbing, music generation, and conversational voice agents for creators and enterprises.
Industry: AI & Machine Learning
Company Size: Large (251 to 1,000 employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series C
Headquarters: London, United Kingdom
Founded: 2022
Glassdoor
Glassdoor: 4.2
WebsiteLinkedInGlassdoor