Lead AI Engineer (FM Hosting, LLM Inference)
New York, Cambridge, San Jose
Workplace: OnsiteFull timeUSD 215,200 - 245,600 annuallyFunction: Data Science & Machine LearningExperience: 4+ yearsSkills: ["Collaboration","Technical problem-solving","Staying current with research","Communicating clearly","Adapting to ambiguity"]Build and deploy foundational AI capabilities across large-scale LLM inference and related components, working with engineers, research scientists, TPMs, and product managers. Design, develop, test, and support AI software for foundation model training, similarity search, guardrails, model evaluation, experimentation, governance, and observability. Use Open Source and SaaS AI technologies on cloud platforms and invent LLM optimization techniques to improve scalability, cost, latency, and throughput.
Loading
Loading job details...
Preparing the role view and application actions.

