Senior Engineer II, Serverless Inference

Digital Ocean
Seattle
Workplace: HybridFull timeUSD 167,000 - 209,000Function: Hospitality & Food ServiceExperience: 7+ yearsSkills: ["Ownership","Accountability","Continuous learning","Technical leadership","Incident response"]

Build and optimize DigitalOcean’s serverless inference infrastructure and APIs to support large-scale AI workloads. You’ll design multi-tenant, production-grade services with intelligent routing, improve resiliency through observability and automation, and partner with platform and GPU infrastructure teams. Lead technical efforts across architecture, traffic management, reliability, and incident response, while coaching engineers and raising engineering standards for continuous improvement.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Digital Ocean
Digital Ocean
16 hours ago

Senior Engineer II, Serverless Inference

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 3 hours agoStatus: Live

Job Summary

Build and optimize DigitalOcean’s serverless inference infrastructure and APIs to support large-scale AI workloads. You’ll design multi-tenant, production-grade services with intelligent routing, improve resiliency through observability and automation, and partner with platform and GPU infrastructure teams. Lead technical efforts across architecture, traffic management, reliability, and incident response, while coaching engineers and raising engineering standards for continuous improvement.
Location: Seattle
Workplace: Hybrid
Employment Type: Full time
Job Function: Hospitality & Food Service
Seniority: Mid level

Key Responsibilities

  • •Design and build scalable, multi-tenant services for AI inference and intelligent routing workloads.
  • •Improve platform resiliency via observability, capacity management, automation, and operational tooling.
  • •Partner with platform, GPU infrastructure, and product engineering teams to deliver production-grade systems and highly available APIs.
  • •Drive architecture decisions across traffic management, service orchestration, reliability, and platform scalability.
  • •Participate in on-call rotations, lead incident response and operational readiness improvements, and coach engineers.

Pay and Benefits

Salary: USD 167,000 - 209,000
Equity and Bonus:Equity
Perks:Employee AssistanceLearning Budget

Key Requirements

  • •7+ years building and operating multi-tenant platforms or distributed backend systems.
  • •Strong experience operating high-scale distributed services in production environments.
  • •Deep SRE knowledge: observability, incident management, reliability engineering, capacity planning, and operational automation.
  • •1+ years hands-on Go/Golang experience in production systems.
  • •2+ years experience with Kubernetes, plus experience debugging performance, scalability, and reliability issues.
Experience:7+ yearsAICloud computingDistributed systemsSREMulti-tenant platformsKubernetesLLM inference
Skills:OwnershipAccountabilityContinuous learningTechnical leadershipIncident response
Languages:English
Tech Stack:GoGolangKubernetesMicroservicesDistributed systemsObservabilityIncident managementReliability engineeringCapacity planningAutomationAPI gatewaysService meshGPU utilizationTime To First Token (TTFT)Time Per Output Token (TPOT)VLLMTritonTensorRT-LLMRate limitingTraffic routing

Company Brief

Digital Ocean
Provides cloud infrastructure and developer-focused cloud services including scalable droplets, managed databases, Kubernetes, object storage, and networking to simplify deploying and managing applications for developers and small-to-medium businesses.
Industry: Cloud Computing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 250M to 500M
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: New York, United States
Founded: 2011
Glassdoor
Glassdoor: 3.8
WebsiteLinkedIn