Research Staff, Data Science

Deepgram
California, Ann Arbor, San Francisco
Workplace: RemoteFull timeUSD 150,000 - 220,000 annuallyFunction: Data Science & Machine LearningSkills: ["Python","PyTorch","Communication","Data pipelines","Deep learning"]

Seasoned data scientist will build and optimize end-to-end data pipelines for speech and language AI, develop advanced audio analytics, collaborate with DataOps and Engineering, create benchmarks and datasets, and communicate findings to internal and external audiences to advance Deepgram’s voice AI foundation models.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Deepgram
Deepgram
9 months ago

Research Staff, Data Science

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 4 hours agoStatus: Live

Job Summary

Seasoned data scientist will build and optimize end-to-end data pipelines for speech and language AI, develop advanced audio analytics, collaborate with DataOps and Engineering, create benchmarks and datasets, and communicate findings to internal and external audiences to advance Deepgram’s voice AI foundation models.
Location: California, Ann Arbor, San Francisco
Workplace: Remote
Employment Type: Full time
Job Function: Data Science & Machine Learning

Key Responsibilities

  • •Drive high performance data acquisition, preparation and synthesis pipelines to generate data for the next generation of speech and language AI foundation models
  • •Develop advanced characterizations of complex conversational audio utilizing a diverse toolkit of signals processing techniques and deep learning models
  • •Collaborate with DataOps and Engineering to create automated systems which scale the ability of human annotators to label high value data and provide critical feedback on model outputs
  • •Build advanced benchmarking methodologies and curated datasets for evaluating conversational voice systems
  • •Document and present results of data experiments and analysis for internal and external audiences

Pay and Benefits

Salary: USD 150,000 - 220,000 annually
Equity and Bonus:Equity
Perks:Health InsuranceDentalVision401kPaid LeaveParential LeaveHome OfficeWellness StipendLife InsuranceLearning Stipend

Key Requirements

  • •Experience building data processing pipelines from a blank page and owning the entire data stack including data acquisition, characterization, cleaning, serving and transformation
  • •Experience and expertise applying statistical methods and deep learning models to understand complex data
  • •Strong communication skills and the ability to translate complex concepts in simple terms, depending on the target audience
  • •Strong software engineering skills with particular emphasis on developing clean, modular code in Python and working with PyTorch
  • •Background in speech and audio data or related domain experience is a plus
Experience:SpeechAudioData science
Skills:PythonPyTorchCommunicationData pipelinesDeep learning
Languages:English
Tech Stack:PythonPyTorchData pipelinesSignals processingAIMachine learning

Company Brief

Deepgram
Deepgram builds a real-time Voice AI platform delivering speech-to-text, text-to-speech, and voice-agent APIs for developers and enterprises, focusing on low-latency, high-accuracy voice models and scalable deployment options.
Industry: API Platforms
Company Size: Medium (51 to 250 employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series C
Headquarters: San Francisco, United States
Founded: 2015
Glassdoor
Glassdoor: 4.5
WebsiteLinkedInGlassdoor