AI Inference Engineer

Baseten
San Francisco, Toronto, New York, Montreal
Workplace: HybridFull timeUSD 165,000 - 330,000 annuallyFunction: Data Science & Machine LearningExperience: 2+ yearsEducation: bachelorsSkills: ["Communication","Ownership","Accountability","Judgment","Collaboration"]

Partner directly with customers to architect, build, and deploy high-scale production AI applications on Baseten’s platform. Own the journey from problem framing and evaluation through production deployment and monitoring, translating ambiguous goals into reliable, observable services. Ship quickly by turning vague objectives into clear specs and PoCs, while optimizing AI/ML projects and contributing to the evolving technical stack. Balance hands-on engineering with product and customer-facing execution.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Baseten
Baseten
2 days ago

AI Inference Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 11 hours agoStatus: Live

Job Summary

Partner directly with customers to architect, build, and deploy high-scale production AI applications on Baseten’s platform. Own the journey from problem framing and evaluation through production deployment and monitoring, translating ambiguous goals into reliable, observable services. Ship quickly by turning vague objectives into clear specs and PoCs, while optimizing AI/ML projects and contributing to the evolving technical stack. Balance hands-on engineering with product and customer-facing execution.
Location: San Francisco, Toronto, New York, Montreal
Workplace: Hybrid
Employment Type: Full time
Job Function: Data Science & Machine Learning
Seniority: Mid level

Key Responsibilities

  • •Develop and maintain production software systems and product features, with a preference for Python for ML-relevant work.
  • •Design, implement, and deploy Baseten solutions end-to-end (problem framing → evaluation → production deployment → monitoring) with customer engineering teams.
  • •Turn vague objectives into clear specs and well-defined PoCs to rapidly ship well-tested services and outcomes.
  • •Optimize and enhance AI/ML projects and contribute to continuous improvement of the technical stack, including features and PRDs.
  • •Own products and customer projects end-to-end, operating across engineering, project management, and product management with user empathy and execution focus.

Pay and Benefits

Salary: USD 165,000 - 330,000 annually
Perks:EquityHealth InsuranceDentalVisionPaid Parental401k

Key Requirements

  • •Bachelor's, Master's, or Ph.D. degree in Computer Science, Engineering, Mathematics, or a related field.
  • •2+ years of professional work experience in a fast-paced, high-growth environment.
  • •Demonstrated production experience with one or more general-purpose programming languages, with a strong preference for Python.
  • •Familiarity with AI/ML pipelines and the lifecycle of ML model development and deployment.
  • •Strong communication skills for complex technical topics; experience building or optimizing AI/ML projects is highly valued.
Experience:2+ yearsAI/MLML startups
Education:Bachelor's in Computer Science, Engineering, Mathematics
Skills:CommunicationOwnershipAccountabilityJudgmentCollaboration
Tech Stack:PythonDockerComfyUIWhisper

Company Brief

Baseten
Baseten provides an inference-first ML infrastructure platform that lets engineering and ML teams deploy, serve, and scale machine-learning models with optimized performance, autoscaling, and GPU-backed hosting for production AI applications. ([crunchbase.com](https://www.crunchbase.com/organization/baseten?utm_source=openai))
Industry: AI & Machine Learning
Company Size: Medium (51 to 250 employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2019
Glassdoor
Glassdoor: 5.0
WebsiteLinkedInGlassdoor