Staff Engineer, Inference Optimizations
Digital Ocean
Boston
Workplace: RemoteFull timeUSD 191,200 - 239,000 annuallyFunction: Software EngineeringExperience: 5+ yearsSkills: ["Technical leadership","Code/design review","Cross-functional collaboration"]Lead performance architecture and deep-dive optimization for DigitalOcean’s AI inference stack, improving throughput and reducing latency across inference engine and GPU kernel layers. Drive benchmarking, kernel-level and attention/memory/precision optimizations, and advanced parallelization for multi-node GPU clusters. Serve as a subject matter expert across NVIDIA/AMD ecosystems (CUDA/ROCm/TensorRT/Triton), guide technical roadmap decisions, and translate hardware limits into shippable, developer-friendly product features.

