Senior Machine Learning Engineer, Infrastructure

Patreon
New York, San Francisco
Workplace: HybridFull timeUSD 212,000 - 318,000 annuallyFunction: Data Science & Machine LearningSkills: ["Communication","Debugging","Documentation","Growth mindset","Code review"]

Build and scale live ML inference infrastructure for discovery and feed relevance systems. Own the end-to-end feature store lifecycle and ensure online/offline feature consistency with high availability. Design observability and validation to detect latency, performance gaps, and production drift, and automate model deployment with reliability testing. Collaborate with product, data engineering, and Trust & Safety to turn requirements into robust infrastructure, and debug production bottlenecks.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Patreon
Patreon
3 days ago

Senior Machine Learning Engineer, Infrastructure

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 14 hours agoStatus: Live

Job Summary

Build and scale live ML inference infrastructure for discovery and feed relevance systems. Own the end-to-end feature store lifecycle and ensure online/offline feature consistency with high availability. Design observability and validation to detect latency, performance gaps, and production drift, and automate model deployment with reliability testing. Collaborate with product, data engineering, and Trust & Safety to turn requirements into robust infrastructure, and debug production bottlenecks.
Location: New York, San Francisco
Workplace: Hybrid
Employment Type: Full time
Job Function: Data Science & Machine Learning
Seniority: Mid level

Key Responsibilities

  • •Architect, scale, and maintain high-throughput, low-latency live inference infrastructure for relevance systems.
  • •Own the end-to-end feature store lifecycle from ingestion and transformation to production serving with high availability and consistent online/offline features.
  • •Design and implement observability, monitoring, and validation frameworks to detect performance gaps, latency spikes, and production drift.
  • •Collaborate with product, data engineering, and Trust & Safety to translate requirements into robust, scalable infrastructure solutions.
  • •Automate model deployment and reliability testing and debug complex relevance systems when issues are identified.

Pay and Benefits

Salary: USD 212,000 - 318,000 annually
Equity and Bonus:Equity
Perks:Health InsuranceFlexible Time401kParental LeaveCommuter BenefitsLearning BudgetPaid LeaveLifestyle Stipends

Key Requirements

  • •Deep experience building, deploying, and maintaining production-grade ML infrastructure at scale, including low-latency live inference pipelines and feature store architectures.
  • •Strong background in distributed systems and backend engineering, with ability to write robust, maintainable code in Python.
  • •Systematic debugging skills for complex, high-throughput systems and performance bottlenecks.
  • •Experience building 0 to 1 infrastructure systems that provide a reliable foundation for teams.
  • •Strong communication skills for creating clear documentation for system architectures and infrastructure strategies.
Skills:CommunicationDebuggingDocumentationGrowth mindsetCode review
Tech Stack:Python

Company Brief

Patreon
Patreon is a membership platform that helps creators build paid communities and monetize their work through subscriptions, enabling artists, podcasters, musicians, writers, and other creators to earn recurring revenue from fans.
Industry: Influencer & Creator Economy
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series F
Headquarters: San Francisco, United States
Founded: 2013
Glassdoor
Glassdoor: 3.8
WebsiteLinkedInGlassdoor