Software Engineer - Voice Model

X AI
Palo Alto
Workplace: OnsiteFull timeUSD 150,000 - 450,000 annuallyFunction: Software EngineeringSkills: ["Communication","Problem-solving","Teamwork"]

Join the Grok Voice Model team to build world-class voice AI with low-latency, multilingual spoken interactions. Lead end-to-end training pipelines, data curation, and evaluation frameworks while collaborating with product teams to deploy scalable voice models in real-time environments across devices.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
X AI
X AI
4 months ago

Software Engineer - Voice Model

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 4 hours agoStatus: Live

Job Summary

Join the Grok Voice Model team to build world-class voice AI with low-latency, multilingual spoken interactions. Lead end-to-end training pipelines, data curation, and evaluation frameworks while collaborating with product teams to deploy scalable voice models in real-time environments across devices.
Location: Palo Alto
Workplace: Onsite
Employment Type: Full time
Job Function: Software Engineering

Key Responsibilities

  • •Design and implement large-scale speech data curation and processing pipelines, including collection of real-world audio, synthetic data generation, and automated annotation workflows to enable high-quality model training and evaluation.
  • •Work on pre-training and post-training of speech-language models with supervised fine-tuning, reinforcement learning, and other techniques to ensure accurate, factual, natural, and multilingual Grok Voice responses.
  • •Build and iterate an evaluation framework covering objective metrics, human studies, content factuality, real-time interaction quality, and experimentation infrastructure to measure and improve performance.
  • •Collaborate with product teams to integrate voice models into applications and real-time environments, define spoken interaction specs, and manage lifecycle from prototype to global-scale deployment for low-latency voice experiences.
  • •Drive cross-team initiatives to improve model quality, efficiency, and reliability in production.

Pay and Benefits

Salary: USD 150,000 - 450,000 annually
Equity and Bonus:Equity
Perks:401kHealth InsuranceVisionDentalRemote WorkPaid Leave

Key Requirements

  • •Python expert with deep proficiency in writing clean, efficient code for AI/ML systems.
  • •Hands-on experience processing large-scale datasets using Spark and Ray for cleaning, augmentation, and feature extraction.
  • •Proficiency in pre-training and post-training speech-language models using JAX/PyTorch, including supervised fine-tuning, reinforcement learning, and optimizations for accuracy, factuality, natural spoken style, detail, and multilingual fluency.
  • •Ability to set up and run rigorous evaluation pipelines: objective metrics, human preference studies, content factuality checks, and iterative A/B testing to drive model improvements.
  • •Experience building or working with large-scale distributed training and inference systems on Kubernetes.
Experience:AISpeechNLPDistributed training
Skills:CommunicationProblem-solvingTeamwork
Languages:English
Tech Stack:PythonSparkRayJAXPyTorchKubernetes

Company Brief

X AI
Develops advanced artificial intelligence models and research aimed at building safe, general AI and understanding the fundamental nature of the universe. Focuses on large-scale AI systems, research publications, and building foundational AI capabilities.
Industry: AI & Machine Learning
Company Size: Medium (51 to 250 employees)
Growth: Early Stage Startup
Headquarters: San Francisco, United States
Founded: 2023
Website