Engineering Manager, Forward Deployed Engineering (LLM)

Baseten
San Francisco, New York, Toronto, Canada, Montreal
Workplace: HybridFull timeUSD 260,000 - 380,000 annuallyFunction: OtherExperience: 4+ yearsEducation: bachelorsSkills: ["Communication","Leadership","Collaboration"]

Lead and mentor a team of Forward Deployed Engineers focused on building, scaling, and optimizing LLM inference workloads for Baseten customers. Combine hands-on technical ownership with people leadership to design, deploy, and manage high-performance, low-latency AI applications on Baseten’s platform, collaborating across product, infrastructure, and customer teams.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Baseten
Baseten
9 months ago

Engineering Manager, Forward Deployed Engineering (LLM)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 10 hours agoStatus: Live

Job Summary

Lead and mentor a team of Forward Deployed Engineers focused on building, scaling, and optimizing LLM inference workloads for Baseten customers. Combine hands-on technical ownership with people leadership to design, deploy, and manage high-performance, low-latency AI applications on Baseten’s platform, collaborating across product, infrastructure, and customer teams.
Location: San Francisco, New York, Toronto, Canada, Montreal
Workplace: Hybrid
Employment Type: Full time
Seniority: Manager level

Key Responsibilities

  • •Lead, mentor, and grow a team of Forward Deployed Engineers, providing guidance on technical direction, project execution, and professional development.
  • •Set clear goals and ensure timely, high-quality delivery across multiple customer-facing projects involving LLM deployment and inference optimization.
  • •Collaborate with leadership to align team priorities with company and customer goals, balancing short-term delivery, widely varying customer priorities, and long-term technical initiatives.
  • •Player-coach – be a hands-on contributor on strategic product initiatives and customer engagements as needed.
  • •Own software systems and product features in a production environment, with a focus on Python and ML inference optimization, from problem framing to production deployment and monitoring.

Pay and Benefits

Salary: USD 260,000 - 380,000 annually
Equity and Bonus:Equity
Perks:Equity401kMedicalDentalVisionPar Parental

Key Requirements

  • •Bachelor’s, Master’s, or Ph.D. in Computer Science, Engineering, or related field.
  • •4+ years of professional software engineering experience, including 1+ year in a leadership or mentorship capacity.
  • •Strong programming skills in Python, with production experience in building or optimizing ML inference systems.
  • •Proven experience with LLMs, inference optimization, or serving frameworks (e.g., vLLM, TensorRT, Triton, Hugging Face, Ray Serve).
  • •Familiarity with observability, profiling, and cost/performance tradeoffs in production ML systems.
Experience:4+ yearsAILLMMachine learning
Education:Bachelor's
Skills:CommunicationLeadershipCollaboration
Tech Stack:PythonLLMsTensorRTTritonHugging FaceRay Serve

Company Brief

Baseten
Baseten provides an inference-first ML infrastructure platform that lets engineering and ML teams deploy, serve, and scale machine-learning models with optimized performance, autoscaling, and GPU-backed hosting for production AI applications. ([crunchbase.com](https://www.crunchbase.com/organization/baseten?utm_source=openai))
Industry: AI & Machine Learning
Company Size: Medium (51 to 250 employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2019
Glassdoor
Glassdoor: 5.0
WebsiteLinkedInGlassdoor