Sr. Lead AI Engineer (Inference Optimization, FM hosting, AI Platform)
Cambridge, New York, San Francisco, San Jose
Workplace: OnsiteFull timeUSD 250,800 - 286,200 annuallyFunction: Data Science & Machine LearningExperience: 6+ yearsEducation: bachelorsSkills: ["Communication","Problem-solving","Leadership","Mentorship","Technical rigor"]Design and deliver AI platform components for large-scale production systems, including foundation model training, LLM inference, similarity search, guardrails, evaluation, experimentation, governance, and observability. Partner with cross-functional engineering, research, technical program, and product teams to build responsible, scalable AI-powered experiences. Leverage AWS Ultraclusters, Hugging Face, VectorDBs, Nemo Guardrails, and PyTorch, and invent optimization techniques to improve scalability, cost, latency, and throughput.
Loading
Loading job details...
Preparing the role view and application actions.

