Senior Forward Deployed Engineer I (AI Inference)
Bengaluru
Workplace: HybridFull timeFunction: Data Science & Machine LearningSkills: ["Growth mindset","Technical leadership","Customer empathy","Communication","Bias for action"]Embed with AI founders and strategic AI enterprises to build and deploy low-latency, high-throughput LLM inference on DigitalOcean’s GPU cloud. Own the full lifecycle of production-grade, multi-tenant inference infrastructure—architecting distributed systems, profiling bottlenecks, and fixing KV-cache locality issues. Lead customer-facing debugging and performance work to reduce time-to-first-token and time-per-output-token using Kubernetes-native tooling and GPU-efficiency techniques.
Loading
Loading job details...
Preparing the role view and application actions.

