Senior Lead AI Engineer (FM Hosting, LLM Inference)
New York, San Jose
Workplace: OnsiteFull timeUSD 229,900 - 262,400 annuallyFunction: Data Science & Machine LearningExperience: 6+ yearsEducation: bachelorsSkills: ["Communication","Presentation","Collaboration"]Design, develop, test, deploy, and support foundational AI software for large-scale LLM inference and responsible AI capabilities. Partner with engineers, research scientists, program managers, and product managers to deliver AI-powered products and build core AI infrastructure. Leverage tools such as AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, and PyTorch, while inventing LLM optimization techniques to improve scalability, cost, latency, and throughput.
Loading
Loading job details...
Preparing the role view and application actions.

