Systems Architect AI/ML Infrastructure

Deepgram
United States
Workplace: RemoteFull timeUSD 160,000 - 220,000 annuallyFunction: Software EngineeringExperience: 7+ yearsSkills: ["Kubernetes","GPU","Storage","FinOps","Multi-cloud","Cloud"]

Senior infrastructure leader responsible for end-to-end AI/ML infrastructure architecture across production inference and research training. Designs multi-cloud/hybrid strategies, compute orchestration for GPU/CPU workloads, scalable storage for large datasets, and capacity planning with cost optimization to support Deepgram’s real-time voice AI platform.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Deepgram
Deepgram
5 months ago

Systems Architect AI/ML Infrastructure

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 15 hours agoStatus: Live

Job Summary

Senior infrastructure leader responsible for end-to-end AI/ML infrastructure architecture across production inference and research training. Designs multi-cloud/hybrid strategies, compute orchestration for GPU/CPU workloads, scalable storage for large datasets, and capacity planning with cost optimization to support Deepgram’s real-time voice AI platform.
Location: United States
Workplace: Remote
Employment Type: Full time
Job Function: Software Engineering

Key Responsibilities

  • •Define and drive the end-to-end infrastructure architecture for Deepgram's AI/ML workloads across production inference and research training
  • •Design multi-cloud and hybrid infrastructure strategies that balance performance, reliability, cost, and vendor flexibility
  • •Architect compute orchestration systems that efficiently schedule and manage GPU and CPU workloads across heterogeneous infrastructure
  • •Design storage architectures that handle the massive datasets required for speech and audio ML -- from high-throughput training data pipelines to low-latency model serving
  • •Lead capacity planning across all infrastructure dimensions, modeling growth and ensuring Deepgram can scale ahead of demand

Pay and Benefits

Salary: USD 160,000 - 220,000 annually
Equity and Bonus:Equity
Perks:Health InsuranceDentalVisionWellness StipendPaid LeaveParential Leave401kHome OfficeLearning Stipend

Key Requirements

  • •7+ years of experience in infrastructure engineering, systems architecture, or a senior technical role focused on large-scale infrastructure
  • •Proven experience designing multi-cloud architectures spanning AWS and at least one other major cloud provider or on-premises environment
  • •Deep expertise in storage system design -- block, object, and file storage, including performance tuning for large-scale data workloads
  • •Strong experience with compute orchestration using Kubernetes, and an understanding of how to schedule diverse workloads efficiently
  • •Hands-on experience with GPU infrastructure -- procurement considerations, cluster design, driver and runtime management
Experience:7+ yearsAI/MLCloudInfrastructure
Skills:KubernetesGPUStorageFinOpsMulti-cloudCloud
Tech Stack:KubernetesGPUCPUStorageMulti-cloudCloudFinOps

Company Brief

Deepgram
Deepgram builds a real-time Voice AI platform delivering speech-to-text, text-to-speech, and voice-agent APIs for developers and enterprises, focusing on low-latency, high-accuracy voice models and scalable deployment options.
Industry: API Platforms
Company Size: Medium (51 to 250 employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series C
Headquarters: San Francisco, United States
Founded: 2015
Glassdoor
Glassdoor: 4.5
WebsiteLinkedInGlassdoor