Staff Applied AI Inference Engineer
Denver
Workplace: OnsiteFull timeUSD 185,000 - 225,000Function: Data Science & Machine LearningSkills: ["Communication","Problem-solving","Opportunity-finding","Urgency","Ownership"]Own the end-to-end AI inference stack to make large language models faster, cheaper, and more reliable in production. Bring modern inference optimizations into real deployments by profiling latency and cost, working down into serving code and CUDA kernels, and tuning for GPU performance. Partner with customer engineering teams to take workloads from proof of concept to monitored production services while keeping targets for throughput, latency, and dependability.
Loading
Loading job details...
Preparing the role view and application actions.

