Senior Staff Applied AI Inference Engineer
San Francisco
Workplace: OnsiteFull timeUSD 250,000 - 300,000 annuallyFunction: Data Science & Machine LearningEducation: bachelorsSkills: ["Problem-solving","Opportunity-finding","Urgency","Communication","Ownership"]Build and own the inference stack end to end to make large language model serving faster, cheaper, and more reliable in production. Profile where time and cost go, apply modern optimization techniques, and work deep into serving code—from vLLM and SGLang down to CUDA kernels. Partner with customer engineering teams to tailor deployments to real models, traffic, latency, and cost constraints, taking workloads from proof of concept to monitored production services.
Loading
Loading job details...
Preparing the role view and application actions.

